Pith. sign in

REVIEW 17 cited by

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2507.14267 v1 pith:3NNIK2NB submitted 2025-07-18 cs.AI cond-mat.mtrl-sci

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

classification cs.AI cond-mat.mtrl-sci
keywords dreamsmaterialssimulationagenticagentscapabilitiesdensitydiscovery
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Materials discovery relies on high-throughput, high-fidelity simulation techniques such as Density Functional Theory (DFT), which require years of training, extensive parameter fine-tuning and systematic error handling. To address these challenges, we introduce the DFT-based Research Engine for Agentic Materials Screening (DREAMS), a hierarchical, multi-agent framework for DFT simulation that combines a central Large Language Model (LLM) planner agent with domain-specific LLM agents for atomistic structure generation, systematic DFT convergence testing, High-Performance Computing (HPC) scheduling, and error handling. In addition, a shared canvas helps the LLM agents to structure their discussions, preserve context and prevent hallucination. We validate DREAMS capabilities on the Sol27LC lattice-constant benchmark, achieving average errors below 1\% compared to the results of human DFT experts. Furthermore, we apply DREAMS to the long-standing CO/Pt(111) adsorption puzzle, demonstrating its long-term and complex problem-solving capabilities. The framework again reproduces expert-level literature adsorption-energy differences. Finally, DREAMS is employed to quantify functional-driven uncertainties with Bayesian ensemble sampling, confirming the Face Centered Cubic (FCC)-site preference at the Generalized Gradient Approximation (GGA) DFT level. In conclusion, DREAMS approaches L3-level automation - autonomous exploration of a defined design space - and significantly reduces the reliance on human expertise and intervention, offering a scalable path toward democratized, high-throughput, high-fidelity computational materials discovery.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 17 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. FermiLink: A Unified Agent Framework for Multidomain Autonomous Scientific Simulations

    physics.chem-ph 2026-04 conditional novelty 8.0

    FermiLink is a unified AI agent framework that automates multidomain scientific simulations via separated package knowledge bases and a four-layer progressive disclosure mechanism, reproducing 56% of target figures in...

  2. Lang2MLIP: End-to-End Language-to-Machine Learning Interatomic Potential Development with Autonomous Agentic Workflows

    cs.LG 2026-05 unverdicted novelty 7.0

    Lang2MLIP is an LLM multi-agent framework that automates end-to-end development of machine learning interatomic potentials from natural language input for heterogeneous materials systems.

  3. El Agente Quntur: A research collaborator agent for quantum chemistry

    physics.chem-ph 2026-02 unverdicted novelty 7.0

    El Agente Quntur is a new multi-agent system that uses reasoning over literature and software documentation to autonomously handle the full workflow of quantum chemistry experiments in ORCA.

  4. Grounded autonomous scrutiny at scale: emergent critique from reproduction of published computational physics papers

    physics.comp-ph 2026-04 unverdicted novelty 6.5

    LLM agents surface methodological critiques of computational-physics papers primarily by re-running calculations, not by reading; a deep case revises a Nature Communications L_G=5 nm claim with attacks missing from 21...

  5. AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

    cond-mat.mtrl-sci 2026-05 unverdicted novelty 6.0

    AutoDFT presents a closed-loop multi-agent LLM framework achieving 94.1% success on a 34-task DFT benchmark and reliable property predictions on materials databases.

  6. TSAgent: An Agentic Workflow for Autonomous Transition State Search

    physics.chem-ph 2026-05 unverdicted novelty 6.0

    TSAgent automates transition state searches at DFT accuracy via an agentic loop, reaching 83% success on 100 OC20NEB examples and 70% on 10 held-out cases versus 73% for human experts.

  7. ChemGraph-XANES: An Agentic Framework for XANES Simulation and Curation

    cond-mat.mtrl-sci 2026-04 unverdicted novelty 6.0

    An LLM-orchestrated framework automates the full XANES workflow from natural language to normalized spectra and curated data.

  8. Grounded autonomous scrutiny at scale: emergent critique from reproduction of published computational physics papers

    physics.comp-ph 2026-04 conditional novelty 6.0

    An LLM agent autonomously runs read-plan-compute-compare loops on 111 computational physics papers, raising substantive concerns in 42% of them (97.7% only after execution), and generates a full publishable Comment re...

  9. Evaluating LLM-generated code for domain-specific languages: molecular dynamics with LAMMPS

    cs.SE 2026-03 unverdicted novelty 6.0

    LLM syntax accuracy for LAMMPS scripts improved to 91% parser pass rate, yet only 1/80 scripts were scientifically correct on the hardest prompt; an agentic verification skill raised success to 5/6.

  10. QUASAR: A Universal Autonomous System for Atomistic Simulation and a Benchmark of Its Capabilities

    cond-mat.mtrl-sci 2026-01 unverdicted novelty 6.0

    QUASAR is a new autonomous LLM-based system that orchestrates multi-scale atomistic simulations and benchmarks as a general reasoning tool rather than a narrow automation script.

  11. Toward Exascale AI for Science: A Scalable AI Skill for Autonomous Microkinetics Discovery

    cs.CE 2026-06 conditional novelty 5.0

    An agentic HPC skill automates NEB microkinetics, recovers from common failures, and benchmarks ~12 universal MLIPs against DFT for CO2 sublimation on graphite.

  12. El Agente Estructural: An Artificially Intelligent Molecular Editor

    physics.chem-ph 2026-02 unverdicted novelty 5.0

    El Agente Estructural is a new multimodal agent that performs natural-language-driven 3D molecular geometry editing and generation using integrated domain tools and vision-language models.

  13. A Robust Agentic Framework for Expert-Level Automation of Atomistic Simulations

    cond-mat.mtrl-sci 2026-06 unverdicted novelty 4.0

    Paimon is an agentic framework that automates atomistic simulations and improves reliability by suppressing silent errors in agent workflows, demonstrated on liquid electrolyte cases and literature reproduction.

  14. LARA: Validation-Driven Agentic Supercomputer Workflows for Atomistic Modeling

    physics.comp-ph 2026-04 unverdicted novelty 4.0

    LARA-HPC introduces a validation-first agentic system with dry-run verification and multi-phase refinement that improves robustness of AI-generated DFT workflows on HPC systems.

  15. ChemGraph-XANES: An Agentic Framework for XANES Simulation and Curation

    cond-mat.mtrl-sci 2026-04 conditional novelty 4.0

    ChemGraph-XANES is an LLM-based agentic framework that automates FDMNES XANES simulation workflows via schema-constrained tool execution and documentation-grounded parameter selection.

  16. RADIANT-LLM: an Agentic Retrieval Augmented Generation Framework for Reliable Decision Support in Safety-Critical Nuclear Engineering

    cs.IR 2026-03 unverdicted novelty 4.0

    RADIANT-LLM is a local-first multi-modal RAG system with provenance tracking that delivers lower hallucination rates than general LLMs on nuclear engineering benchmarks.

  17. Toward Exascale AI for Science: A Scalable AI Skill for Autonomous Microkinetics Discovery

    cs.CE 2026-06 unverdicted novelty 3.0

    Introduces a scalable AI skill framework for autonomous microkinetics discovery that automates workflows and evaluates surrogate reliability.