Pith. sign in

REVIEW 5 cited by

Towards Fast, Specialized Machine Learning Force Fields: Distilling Foundation Models via Energy Hessians

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2501.09009 v2 pith:5L3ATRCJ submitted 2025-01-15 physics.chem-ph cond-mat.mtrl-scics.LGphysics.bio-ph

Towards Fast, Specialized Machine Learning Force Fields: Distilling Foundation Models via Energy Hessians

classification physics.chem-ph cond-mat.mtrl-scics.LGphysics.bio-ph
keywords foundationmodelmodelsspecializedenergymlffmlffswhile
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The foundation model (FM) paradigm is transforming Machine Learning Force Fields (MLFFs), leveraging general-purpose representations and scalable training to perform a variety of computational chemistry tasks. Although MLFF FMs have begun to close the accuracy gap relative to first-principles methods, there is still a strong need for faster inference speed. Additionally, while research is increasingly focused on general-purpose models which transfer across chemical space, practitioners typically only study a small subset of systems at a given time. This underscores the need for fast, specialized MLFFs relevant to specific downstream applications, which preserve test-time physical soundness while maintaining train-time scalability. In this work, we introduce a method for transferring general-purpose representations from MLFF foundation models to smaller, faster MLFFs specialized to specific regions of chemical space. We formulate our approach as a knowledge distillation procedure, where the smaller "student" MLFF is trained to match the Hessians of the energy predictions of the "teacher" foundation model. Our specialized MLFFs can be up to 20 $\times$ faster than the original foundation model, while retaining, and in some cases exceeding, its performance and that of undistilled models. We also show that distilling from a teacher model with a direct force parameterization into a student model trained with conservative forces (i.e., computed as derivatives of the potential energy) successfully leverages the representations from the large-scale teacher for improved accuracy, while maintaining energy conservation during test-time molecular dynamics simulations. More broadly, our work suggests a new paradigm for MLFF development, in which foundation models are released along with smaller, specialized simulation "engines" for common chemical subsets.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Lang2MLIP: End-to-End Language-to-Machine Learning Interatomic Potential Development with Autonomous Agentic Workflows

    cs.LG 2026-05 unverdicted novelty 7.0

    Lang2MLIP is an LLM multi-agent framework that automates end-to-end development of machine learning interatomic potentials from natural language input for heterogeneous materials systems.

  2. Distilling first-principles accuracy into compact machine learning potentials for condensed-phase chemistry

    physics.chem-ph 2026-06 unverdicted novelty 6.0

    Distilled compact MLIPs from transfer-learned teachers reproduce observables more reliably than same-size models trained directly and enable practical PIMD umbrella sampling of water dissociation at TiO2 interface wit...

  3. Compact SO(3) Equivariant Atomistic Foundation Models via Structural Pruning

    cs.LG 2026-05 unverdicted novelty 6.0

    Structural pruning of SO(3) equivariant atomistic models from large checkpoints yields 1.5-4x fewer parameters and 2.5-4x less pre-training compute than small models trained from scratch, while outperforming them on m...

  4. A Lightweight Universal Machine-Learning Interatomic Potential via Knowledge Distillation for Scalable Atomistic Simulations

    cond-mat.mtrl-sci 2026-04 unverdicted novelty 5.0

    SevenNet-Nano is a lightweight universal ML interatomic potential distilled from a larger multi-task foundation model, delivering high accuracy, transferability, and over 10x computational speedup for scalable atomist...

  5. Machine Learning Interatomic Potentials: Advancing Open-Source Software for Efficient and Scalable Molecular Simulation

    physics.chem-ph 2026-05 unverdicted novelty 4.0

    mlip v2 is a new software release that integrates API redesign, e3j backend, eSEN model, improved charge modeling, and expanded simulation capabilities to support larger-scale molecular modeling.