hep-phSep 29, 2026

Searching for BSM Experimental Signatures with Large Lagrangian Models

Authors: Ibrahim Elsharkawy, Victoria Knapp-Perez, Wahid Bhimji, Aishik Ghosh

Organizations: Department of Physics, University of Toronto and Vector Institute, Toronto, ON, Canada · NERSC, Lawrence Berkeley National Laboratory, Berkeley, California, USA · Department of Physics and Astronomy, University of California, Irvine, CA 92697 · Halluminate, San Francisco, California, USA, 94107 · Georgia Institute of Technology, Atlanta, GA 30332 · Lawrence Berkeley National Laboratory, Berkeley, CA 94720

Abstract

The search for physics Beyond the Standard Model (BSM) is generally limited not by the supply of theory descriptions but by the lack of discriminating experimental observations. A case in point is dark matter, where the overwhelming gravitational evidence only goes so far in distinguishing between models within a vast theory space. Exploring the space of testable model signatures may help identify overlooked experimental observables and indicate the utility of future experiments. A challenge is designing a search through model signatures outside what is found in the literature. Our primary contribution is hAIthem, a framework that combines the self-guided exploration of reinforcement learning (RL) with the broad literature-derived knowledge of LLMs. We build an RL agent that learns to find which portions of a theory's high-dimensional parameter space are not excluded under some subset of constraints by playing a Battleship-style "game" against a suite of phenomenology tools. The agent is built as a Large Lagrangian Model (LLaM), an autoregressive transformer that reads a tokenized Lagrangian, is pretrained at scale (here on ~1 billion tokens from ~10,000 Lagrangians), and is fine-tuned in a live environment. The framework then constructs a decision tree that separates RL-found regions using observables computed with established tools, and passes the remaining degenerate regions to a set of LLM agents that compete to produce realistic signatures. In this proof of concept, RL-search outperforms an evolutionary-algorithm baseline, finding more viable regions with greater physical diversity. In a restricted space of single dark scalar multiplet models, we find that hAIthem proposes interesting combinations of previously studied observables, such as the application of a halo-independent kinematic ratio to paleo-detectors.

Figures & tables

Appendix figures & tables27 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Large Discovery Models: Empirically-grounded Model-Based Open-Ended Search

    Aug 16, 2026Zhongwei Yu, Yan Song, Xue Yan +9Model DiscoveryDiscovery

  2. A Probabilistic Framework for LLM-Based Model Discovery

    Feb 20, 2026Stefan Wahl, Raphaela Schenk, Ali Farnoud +2Model DiscoveryProbabilistic Model