REVIEW 4 cited by
Retrieve-Plan-Generation: An Iterative Planning and Answering Framework for Knowledge-Intensive LLM Generation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Despite the significant progress of large language models (LLMs) in various tasks, they often produce factual errors due to their limited internal knowledge. Retrieval-Augmented Generation (RAG), which enhances LLMs with external knowledge sources, offers a promising solution. However, these methods can be misled by irrelevant paragraphs in retrieved documents. Due to the inherent uncertainty in LLM generation, inputting the entire document may introduce off-topic information, causing the model to deviate from the central topic and affecting the relevance of the generated content. To address these issues, we propose the Retrieve-Plan-Generation (RPG) framework. RPG generates plan tokens to guide subsequent generation in the plan stage. In the answer stage, the model selects relevant fine-grained paragraphs based on the plan and uses them for further answer generation. This plan-answer process is repeated iteratively until completion, enhancing generation relevance by focusing on specific topics. To implement this framework efficiently, we utilize a simple but effective multi-task prompt-tuning method, enabling the existing LLMs to handle both planning and answering. We comprehensively compare RPG with baselines across 5 knowledge-intensive generation tasks, demonstrating the effectiveness of our approach.
Forward citations
Cited by 4 Pith papers
-
Condition-Gated Reasoning for Context-Dependent Biomedical Question Answering
A condition-gated knowledge-graph method improves biomedical QA when patient-specific contraindications change the correct answer, and a new 100-question benchmark measures this capability.
-
ScalingNote: Scaling up Retrievers with Large Language Models for Real-World Dense Retrieval
Train dual LLM towers for dense retrieval, then distill the query tower into a small BERT encoder, keeping most of the accuracy gain without the online latency.
-
Large Language Models for Planning: A Comprehensive and Systematic Survey
A structured survey of LLM planning methods, benchmarks, and interpretability work, organized around a three-way taxonomy.
-
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
A systems architecture for SLA-driven reconfiguration of multi-agent RAG is described, but the central reconfiguration claim is not validated by experiments.
Discussion (0). Continue with ORCID to comment.