REVIEW 15 cited by
CRISPR-GPT for Agentic Automation of Gene-editing Experiments
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The introduction of genome engineering technology has transformed biomedical research, making it possible to make precise changes to genetic information. However, creating an efficient gene-editing system requires a deep understanding of CRISPR technology, and the complex experimental systems under investigation. While Large Language Models (LLMs) have shown promise in various tasks, they often lack specific knowledge and struggle to accurately solve biological design problems. In this work, we introduce CRISPR-GPT, an LLM agent augmented with domain knowledge and external tools to automate and enhance the design process of CRISPR-based gene-editing experiments. CRISPR-GPT leverages the reasoning ability of LLMs to facilitate the process of selecting CRISPR systems, designing guide RNAs, recommending cellular delivery methods, drafting protocols, and designing validation experiments to confirm editing outcomes. We showcase the potential of CRISPR-GPT for assisting non-expert researchers with gene-editing experiments from scratch and validate the agent's effectiveness in a real-world use case. Furthermore, we explore the ethical and regulatory considerations associated with automated gene-editing design, highlighting the need for responsible and transparent use of these tools. Our work aims to bridge the gap between beginner biological researchers and CRISPR genome engineering techniques, and demonstrate the potential of LLM agents in facilitating complex biological discovery tasks. The published version of this draft is available at https://www.nature.com/articles/s41551-025-01463-z.
Forward citations
Cited by 15 Pith papers
-
TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability
The paper introduces TCS-Bench, a 300-task proof-generation benchmark from top TCS papers, and reports frontier LLM accuracies from 30% to 68% using an automated verifier.
-
On Path to Multimodal Historical Reasoning: HistBench and HistAgent
HistAgent, a history-specialized agent, scores 27.54% pass@1 and 36.47% pass@2 on the new 414-question HistBench benchmark, surpassing generalist agents tested on the same data.
-
BioMARS: A Multi-Agent Robotic System for Autonomous Biological Experiments
A three-agent LLM/VLM system generated and executed cell-culture protocols on a dual-arm robot, matching manual passaging in viability, with optimization results shown only in a simulated benchmark.
-
Deep Active Learning based Experimental Design to Uncover Synergistic Genetic Interactions for Host Targeted Therapeutics
An ensemble deep active learning framework with knowledge graph embeddings finds 92% of the top 400 HIV double-knockdown pairs after observing less than 6.3% of a 356 by 356 interaction matrix.
-
Aviary: training language agents on challenging scientific tasks
A small open-source LLM trained in the new Aviary environments with expert iteration and majority voting matches or exceeds a frontier LLM agent on SeqQA and LitQA2 at far lower inference cost.
-
MetaScientist: A Human-AI Synergistic Framework for Automated Mechanical Metamaterial Design
A human-in-the-loop LLM and 3D diffusion pipeline generates metamaterial hypotheses and lattice structures, with human-rated novelty but no demonstrated accuracy of target mechanical properties.
-
Toward Greater Autonomy in Materials Discovery Agents: Unifying Planning, Physics, and Scientists
MAPPS combines LLM workflow planning, code generation, and human intuition with machine-learned force fields to discover crystal structures, reporting high stability and novelty rates on MP-20 and Matbench.
-
OpenFOAMGPT 2.0: end-to-end, trustworthy automation for computational fluid dynamics
A multi-agent LLM system converts natural language queries into OpenFOAM simulations and reports 100 percent completion and reproducibility across 455 test cases.
-
LLM4SR: A Survey on Large Language Models for Scientific Research
A systematic review of LLM-based systems for hypothesis discovery, experiment planning, scientific writing, and peer review, including benchmarks, evaluation methods, and open challenges.
-
AI for Auto-Research: Roadmap & User Guide
The paper delivers a stage-by-stage roadmap for AI in research, showing reliable assistance in retrieval and tool tasks but fragility in novelty and judgment, advocating human-governed collaboration.
-
SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration
SciToolAgent uses a knowledge graph of over 500 scientific tools to help LLMs select and chain tools, reaching 94% accuracy on a new 531-question benchmark.
-
ChemAU: Harness the Reasoning of LLMs in Chemical Research with Adaptive Uncertainty Estimation
ChemAU adds a position penalty to token-level uncertainty estimates so that flagged reasoning steps are corrected by a fine-tuned chemistry model, reporting improved accuracy on GPQA, MMLU-Pro, and SuperGPQA chemistry...
-
A Survey of Large Language Models in Discipline-specific Research: Challenges, Methods and Opportunities
A review that categorizes methods for adapting LLMs to discipline-specific research and surveys applications across five broad academic fields.
-
Scientific Hypothesis Generation and Validation: Methods, Datasets, and Future Directions
A survey of LLM-based hypothesis generation and validation whose taxonomy is useful in outline but whose citations and tool descriptions are unreliable.
-
Towards Scientific Discovery with Generative AI: Progress, Opportunities, and Challenges
A position paper proposing a research agenda for AI-driven scientific discovery, centered on benchmarks, science agents, multimodal representations, and unified reasoning.
Discussion (0). Continue with ORCID to comment.