Pith. sign in

REVIEW 4 cited by

Xiwu: A Basis Flexible and Learnable LLM for High Energy Physics

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.08001 v1 pith:RFHRXTS4 submitted 2024-04-08 hep-ph cs.AIcs.CLcs.LGhep-exphysics.comp-ph

classification hep-phcs.AIcs.CLcs.LGhep-exphysics.comp-ph
keywords modelmodelsxiwuapplyingdevelopeddomainfieldfoundation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large Language Models (LLMs) are undergoing a period of rapid updates and changes, with state-of-the-art (SOTA) model frequently being replaced. When applying LLMs to a specific scientific field, it's challenging to acquire unique domain knowledge while keeping the model itself advanced. To address this challenge, a sophisticated large language model system named as Xiwu has been developed, allowing you switch between the most advanced foundation models and quickly teach the model domain knowledge. In this work, we will report on the best practices for applying LLMs in the field of high-energy physics (HEP), including: a seed fission technology is proposed and some data collection and cleaning tools are developed to quickly obtain domain AI-Ready dataset; a just-in-time learning system is implemented based on the vector store technology; an on-the-fly fine-tuning system has been developed to facilitate rapid training under a specified foundation model. The results show that Xiwu can smoothly switch between foundation models such as LLaMA, Vicuna, ChatGLM and Grok-1. The trained Xiwu model is significantly outperformed the benchmark model on the HEP knowledge question-and-answering and code generation. This strategy significantly enhances the potential for growth of our model's performance, with the hope of surpassing GPT-4 as it evolves with the development of open-source models. This work provides a customized LLM for the field of HEP, while also offering references for applying LLM to other fields, the corresponding codes are available on Github.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CLVisc Agent for autonomous relativistic hydrodynamics studies

    nucl-th 2026-07 conditional novelty 6.0 of 10

    An LLM agent autonomously created a CLVisc skill and ran two hydrodynamic studies, finding that the high-temperature branch of η/s dominates flow suppression and that PGCM-uniform 16O decouples ellipticity from size.

  2. A Statistical Physics of Language Model Reasoning

    cs.AI 2025-06 conditional novelty 5.0 of 10

    A switching linear dynamical system on a 40-dimensional projection of LLM hidden states captures about half the variance of reasoning trajectories and predicts belief shifts during adversarial prompts.

  3. Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough

    physics.data-an 2026-07 accept novelty 4.0 of 10

    Verification of ML in fundamental physics is essential precisely when models enter statistical modeling, inference, or hypothesis testing, and is bounded by unavoidable inductive bias, sample complexity, and experimen...

  4. HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency

    hep-ph 2025-12 conditional novelty 4.0 of 10

    HEPTAPOD uses LLM agents to drive FeynRules, MadGraph, Pythia, and analysis tools through schema-validated tool calls and run-card templates, demonstrated on a leptoquark signal scan.

Pith tools