X-03Retrieval & RAG
Cast a wide net with similarity search, then rerank hard, so only genuinely relevant content spends context budget.
Further reading
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks Lewis et al., NeurIPS 2020
- Lost in the Middle: How Language Models Use Long Contexts Liu et al., TACL 2024