REVIEW 3 cited by
Web Retrieval Agents for Evidence-Based Misinformation Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper develops an agent-based automated fact-checking approach for detecting misinformation. We demonstrate that combining a powerful LLM agent, which does not have access to the internet for searches, with an online web search agent yields better results than when each tool is used independently. Our approach is robust across multiple models, outperforming alternatives and increasing the macro F1 of misinformation detection by as much as 20 percent compared to LLMs without search. We also conduct extensive analyses on the sources our system leverages and their biases, decisions in the construction of the system like the search tool and the knowledge base, the type of evidence needed and its impact on the results, and other parts of the overall process. By combining strong performance with in-depth understanding, we hope to provide building blocks for future search-enabled misinformation mitigation systems.
Forward citations
Cited by 3 Pith papers
-
Source or It Didn't Happen: A Multi-Agent Framework for Citation Hallucination Detection
CiteTracer detects hallucinated citations with a 12-code field-level taxonomy and a cascading multi-agent retrieval pipeline, reporting 97.1% accuracy on a synthetic benchmark.
-
CrediBench: Building Web-Scale Network Datasets for Information Integrity
CrediBench presents a one-month, 1-billion-edge Common Crawl web graph with text and 11.5K expert credibility labels, while the abstract's promised 8-month dataset and 85%-accuracy classifier are absent from the paper.
-
Veracity: An Open-Source AI Fact-Checking System
Veracity is an open-source LLM-plus-web-search fact-checking app with a 0 to 100 reliability score and explanations, but no evaluation of its accuracy is included.
Discussion (0). Sign in to comment.