Problem
Semantic similarity alone misses exact identifiers and repository-specific language.
Retrieval quality is an evaluation problem before it is a generation problem.
Hybrid retrieval over 10k+ code chunks combining BM25, FAISS, and cross-encoder reranking with sub-second retrieval.
Exact symbols and lexical matches
Interactive system walkthrough · simulated frontend data · not live telemetry
Semantic similarity alone misses exact identifiers and repository-specific language.
Merge lexical and dense candidates, rerank with a cross-encoder, then evaluate the retrieved context.
Measured precision, recall, F1, keyword hit-rate, and latency against a semantic-only baseline.
Dense retrieval underweighted exact code symbols and project vocabulary.
Use complementary retrievers and evaluate retrieval independently from generation.
30% retrieval F1 improvement versus a semantic-only baseline.