Pith. sign in

REVIEW 3 cited by

Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.16833 v2 pith:JS6QYCYF submitted 2024-07-23 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords llmslong-contextaugmentedcomprehensivecontextscostgenerationhowever
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Retrieval Augmented Generation (RAG) has been a powerful tool for Large Language Models (LLMs) to efficiently process overly lengthy contexts. However, recent LLMs like Gemini-1.5 and GPT-4 show exceptional capabilities to understand long contexts directly. We conduct a comprehensive comparison between RAG and long-context (LC) LLMs, aiming to leverage the strengths of both. We benchmark RAG and LC across various public datasets using three latest LLMs. Results reveal that when resourced sufficiently, LC consistently outperforms RAG in terms of average performance. However, RAG's significantly lower cost remains a distinct advantage. Based on this observation, we propose Self-Route, a simple yet effective method that routes queries to RAG or LC based on model self-reflection. Self-Route significantly reduces the computation cost while maintaining a comparable performance to LC. Our findings provide a guideline for long-context applications of LLMs using RAG and LC.

Discussion (0). Sign in to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Exploring Robust Multi-Agent Workflows for Environmental Data Management

    cs.AI 2026-04 conditional novelty 6.0 of 10

    A role-separated multi-agent workflow with deterministic validation gates blocked a coordinate-transformation error before publication and completed a 2,452-station dataset release in two days, based on two non-contro...

  2. LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework

    cs.CL 2025-07 conditional novelty 6.0 of 10

    LOOM-Scope is a framework that standardizes long-context LLM evaluation across 22 benchmarks and integrates a lightweight 12-benchmark suite, LOOMBench, for fast comprehensive assessment.

  3. QwenLong-CPRS: Towards $\infty$-LLMs with Dynamic Context Optimization

    cs.CL 2025-05 conditional novelty 6.0 of 10

    QwenLong-CPRS is a 7B instruction-guided compressor that shrinks long contexts to query-relevant spans, boosting downstream LLM accuracy and cutting prefill cost.

Pith tools