AgentRivet: an automated system for producing Rivet routines from journal publications
hep-ex, cs.AI, hep-ph
Submitted: 2026-06-11
Updated: 2026-09-14
License: http://creativecommons.org/licenses/by/4.0/
The gist: Particle physics collider experiments provide Rivet routines as part of the analysis preservation strategy for model-independent measurements.
Terminology
Abstract
Particle physics collider experiments provide Rivet routines as part of the analysis preservation strategy for model-independent measurements. Rivet is a C++ toolkit that allow new theoretical models to be compared to the measurements, thus aiding the development and tuning of Monte Carlo event generators as well as searches for physics beyond the Standard Model. However, analysis coverage is known to be incomplete, with only 39% of measurements having documented and publicly available Rivet routines. In this article, we design and implement an automated workflow based on Large Language Models with the goal of providing the missing routines. This multi-step workflow, referred to as AgentRivet, extracts the physics analysis information from published papers and writes the missing Rivet routines, with intermediate code- and physics- reviews as part of an autonomous quality control. We report the results obtained using commercial Large Language Models, provided by OpenAI, Anthropic, and Google, for two recent measurements from the ATLAS and CMS experiments. We find that AgentRivet produces competent Rivet routines with few syntax errors. The physics fidelity of the routines is reasonable and follows the explanations given in the relevant publications. Nevertheless, physics-implementation issues do arise and are investigated using the artefacts produced by AgentRivet. The majority of physics implementation issues arise from subtle-but-ambiguous definitions in the given publication, although some models struggle to implement complex observables even when clear definitions are given.
Sources
- HEPData: a repository for high energy physics data
- Reinterpretation and preservation of data and analyses in HEP
- Recommendations for Best Practices for Data Preservation and Open Science in HEP
- Attention Is All You Need
- Emergent Abilities of Large Language Models
- Are Emergent Abilities of Large Language Models a Mirage?
- News Summarization and Evaluation in the Era of GPT-3
- Evaluating Large Language Models Trained on Code
- AI Agents Can Already Autonomously Perform Experimental High Energy Physics
- Agentic AI -- Physicist Collaboration in Experimental Particle Physics: A Proof-of-Concept Measurement with LEP Open Data
- HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency
- GRACE: an Agentic AI for Particle Physics Experiment Design and Simulation
- Automating High Energy Physics Data Analysis with LLM-Powered Agents
- MadAgents
- Measurements of $ZZ \rightarrow \ell\ell\nu\nu$ and $ZZjj \rightarrow \ell\ell\nu\nu jj$ productions in $pp$ collisions at $\sqrt{s}=13$ TeV with the ATLAS detector
- Precise measurement of the $t\bar{t}$ production cross-section and lepton differential distributions in $e\mu$ dilepton events from $\sqrt{s}=13$ TeV $pp$ collisions with the ATLAS detector
- Measurement and interpretation of inclusive $W\gamma$ production in proton-proton collisions at $\sqrt{s}=13$ TeV using the ATLAS detector
- Measurement of event shape variables using charged particles inside jets in proton-proton collisions at $\sqrt{s}$ = 13 TeV
- The automated computation of tree-level and next-to-leading order differential cross sections, and their matching to parton shower simulations
- A comprehensive guide to the physics and usage of PYTHIA 8.3