Pith. sign in

REVIEW 14 cited by

LLMatDesign: Autonomous Materials Discovery with Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.13163 v1 pith:N7FZQXV3 submitted 2024-06-19 cond-mat.mtrl-sci cs.AIcs.CL

LLMatDesign: Autonomous Materials Discovery with Large Language Models

classification cond-mat.mtrl-sci cs.AIcs.CL
keywords materialsllmatdesigndiscoverylargeautonomouschemicaldatadesign
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Discovering new materials can have significant scientific and technological implications but remains a challenging problem today due to the enormity of the chemical space. Recent advances in machine learning have enabled data-driven methods to rapidly screen or generate promising materials, but these methods still depend heavily on very large quantities of training data and often lack the flexibility and chemical understanding often desired in materials discovery. We introduce LLMatDesign, a novel language-based framework for interpretable materials design powered by large language models (LLMs). LLMatDesign utilizes LLM agents to translate human instructions, apply modifications to materials, and evaluate outcomes using provided tools. By incorporating self-reflection on its previous decisions, LLMatDesign adapts rapidly to new tasks and conditions in a zero-shot manner. A systematic evaluation of LLMatDesign on several materials design tasks, in silico, validates LLMatDesign's effectiveness in developing new materials with user-defined target properties in the small data regime. Our framework demonstrates the remarkable potential of autonomous LLM-guided materials discovery in the computational setting and towards self-driving laboratories in the future.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 14 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. LLM-Guided Test-Time Discovery of Quantum-Chemical Approximation Algorithms

    physics.chem-ph 2026-06 unverdicted novelty 7.0

    LADeQ is an LLM-driven workflow that autonomously discovers and implements approximation algorithms for CCSD and CISD calculations, delivering speedups while respecting user-specified error tolerances.

  2. Matter to Mechanism: A Benchmark for AI Co-Scientists in Materials and Battery Research

    cs.CE 2026-06 unverdicted novelty 7.0

    Introduces the Matter to Mechanism benchmark of 2,645 structured instances and a composite metric suite for evaluating AI co-scientists on problem-to-hypothesis reasoning in battery materials research.

  3. ChatMOSP: A Chemistry-Grounded Mobile Agent for Working-State Catalyst Simulations

    cond-mat.mtrl-sci 2026-05 unverdicted novelty 7.0

    ChatMOSP is an AI agent that maps natural-language descriptions of catalyst environments to validated multiscale simulations of working-state nanoparticle morphology and activity.

  4. LLM-driven design of physics-constrained constitutive models: two agents are better than one

    cs.LG 2026-05 unverdicted novelty 7.0

    A Creator-Inspector multi-agent LLM pipeline for constitutive artificial neural networks increases the rate of models satisfying all nine physical constraints to 100% or 56% depending on the LLM backbone.

  5. El Agente Quntur: A research collaborator agent for quantum chemistry

    physics.chem-ph 2026-02 unverdicted novelty 7.0

    El Agente Quntur is a new multi-agent system that uses reasoning over literature and software documentation to autonomously handle the full workflow of quantum chemistry experiments in ORCA.

  6. AlphaEvolve: A coding agent for scientific and algorithmic discovery

    cs.AI 2025-06 unverdicted novelty 7.0

    AlphaEvolve is an LLM-orchestrated evolutionary coding agent that discovered a 4x4 complex matrix multiplication algorithm using 48 scalar multiplications, the first improvement over Strassen's algorithm in 56 years, ...

  7. Toward Auditable AI Scientists: A Hypothesis Evolution Protocol for LLM Agents

    cs.AI 2026-07 conditional novelty 6.0

    HEP externalizes hypothesis generation, evidence-driven belief updates, and lifecycle verdicts so LLM agents run an auditable hypothesis-test-evidence-belief cycle on materials research questions.

  8. General-purpose LLMs as Constrained Crystal Composition Generators

    cond-mat.mtrl-sci 2026-05 unverdicted novelty 6.0

    General-purpose LLMs recover 96% of low-energy Elpasolites via iterative in-context learning, surpassing task-specific models on an established benchmark.

  9. GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

    cs.AI 2025-07 unverdicted novelty 6.0

    GenoMAS deploys six specialized LLM agents with guided planning to preprocess transcriptomic data and identify genes, reaching 89.13% composite similarity and 60.48% F1 on the GenoTEX benchmark while outperforming pri...

  10. My Chemical Harness: Evolutionary Molecular Design over Synthetic Pathways with Large Language Model Agents

    physics.chem-ph 2026-06 unverdicted novelty 5.0

    My Chemical Harness performs evolutionary molecular design by searching over validated synthetic routes with LLMs restricted to high-level preferences, outperforming baselines on an sEH proxy task across multiple metrics.

  11. PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

    cs.AI 2026-05 unverdicted novelty 5.0

    PRISMat generates crystal slabs with mean absolute errors of 0.188 eV/A² for cleavage energy and 2.79 eV for work function, reducing error by 4× versus the next best model while using less inference time.

  12. Scale-Dependent Input Representation and Confidence Estimation for LLMs in Materials Property Prediction

    cond-mat.mtrl-sci 2026-05 conditional novelty 5.0

    Larger LLMs handle detailed crystal descriptions better than small ones, and mean negative log-likelihood of predicted numbers tracks prediction error after fine-tuning.

  13. Design Topological Materials by Reinforcement Fine-Tuned Generative Model

    cond-mat.mtrl-sci 2025-04 unverdicted novelty 5.0

    Reinforcement fine-tuning of a generative model produces new topological insulators and crystalline insulators, exemplified by Ge2Bi2O6 with a 0.26 eV full band gap.

  14. Evolving Roles of LLMs in Scientific Innovation: Assistant, Collaborator, Scientist, and Evaluator

    cs.DL 2025-07 unverdicted novelty 4.0

    The paper proposes a four-role framework for LLMs in scientific innovation and reviews methods, benchmarks, and limitations across Assistant, Collaborator, Scientist, and Evaluator roles.