Pith. sign in

REVIEW 1 cited by

GeneAgent: Self-verification Language Agent for Gene Set Knowledge Discovery using Domain Databases

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.16205 v1 pith:F2Y6GGZ7 submitted 2024-05-25 cs.AI cs.CL

classification cs.AIcs.CL
keywords genegeneagentknowledgediscoverylanguageself-verificationagentdatabases
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Gene set knowledge discovery is essential for advancing human functional genomics. Recent studies have shown promising performance by harnessing the power of Large Language Models (LLMs) on this task. Nonetheless, their results are subject to several limitations common in LLMs such as hallucinations. In response, we present GeneAgent, a first-of-its-kind language agent featuring self-verification capability. It autonomously interacts with various biological databases and leverages relevant domain knowledge to improve accuracy and reduce hallucination occurrences. Benchmarking on 1,106 gene sets from different sources, GeneAgent consistently outperforms standard GPT-4 by a significant margin. Moreover, a detailed manual review confirms the effectiveness of the self-verification module in minimizing hallucinations and generating more reliable analytical narratives. To demonstrate its practical utility, we apply GeneAgent to seven novel gene sets derived from mouse B2905 melanoma cell lines, with expert evaluations showing that GeneAgent offers novel insights into gene functions and subsequently expedites knowledge discovery.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Knowledge-guided Contextual Gene Set Analysis Using Large Language Models

    q-bio.GN 2025-06 conditional novelty 6.0 of 10

    cGSA combines gene clustering, enrichment analysis, and LLM filtering to prioritize contextually relevant pathways, beating g:Profiler, GPT-4o, and Llama baselines by over 30% on a new manually curated benchmark.

Pith tools