🚀NEW LABGetting Started with Claude AgentsStart lab
Retrieval

EfficientRAG

First page
EfficientRAG
Paper summary

trains an auto-encoder LM to label and tag chunks; it retrieves relevant chunks, tags them as either <Terminate> or <Continue>, and annotates <Continue> chunks for continuous processing; then a filter model is trained to formulate the next-hop query based on the original question and previous annotations; this is done iteratively until all chunks are tagged as <Terminate> or the maximum # of iterations is reached; after the process above has gathered enough information to answer the initial question, the final generator (an LLM) generates the final answer.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack