Glossary · Retrieval engineering
Semantic cache
Reusing an earlier result when a new request is judged meaningfully similar. It reduces latency but can serve stale or inapplicable answers to superficially related questions.
Retrieval engineering
How pages become chunks, vectors and candidates inside an answer engine, and why being retrieved is not being cited.
Corpus · Document ingestion · Data connector · Parsing · OCR · Layout-aware parsing · Semantic chunking · Fixed-size chunking · Chunk overlap · Parent-child retrieval · Metadata filtering · Vector database · Vector index · Sparse vector · Dense vector · Cosine similarity · Dot-product similarity · Approximate nearest neighbor · HNSW · Top-k · Retrieval recall · Retrieval precision · Mean reciprocal rank · Normalized discounted cumulative gain · Relevance score · Candidate generation · Query rewriting · Query expansion · HyDE · Multi-query retrieval · Reciprocal rank fusion · Late interaction · Lexical retrieval · Semantic search · Exact-match retrieval · Freshness boosting · Authority weighting · Source diversity · Deduplication · Near-duplicate detection · Retrieval latency · Retrieval cutoff · Citation mapping · Answer synthesis
← Near-duplicate detection · Retrieval latency →
See it in the full glossary · 579 terms across 19 areas. Scan a site to see which of these apply to it.