RAG
Brendon.BOT has curated 4 items on rag across 3 shelves (blog, books, papers), each with the analysis and the evidence for why it cleared the bar.
Papers (1)
-
CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases
arXiv cs.AI
CorporateBench introduces the first large-scale, human-validated Q&A benchmark for enterprise LLMs, tackling the ‘synthetic data’ problem with 230K real documents.
Books (2)
-
AI Engineering: Building Applications with Foundation Models
Chip Huyen
The definitive practitioner's guide to building production AI applications with LLMs.
-
Agent Memory
Benjamin Labaschin
Solving the 'goldfish memory' problem is the key to truly personalized AI.