RAG vs. long context: choose by failure mode, not by benchmark
Retrieval and million-token windows both promise to put your documents in front of the model. They fail differently, and the failure you can tolerate should decide.
Topic
Retrieval and million-token windows both promise to put your documents in front of the model. They fail differently, and the failure you can tolerate should decide.