Abstract

Much of the recent progress in machine learning domains such as language models has come from scaling laws that predict performance as a function of training effort. In high-energy physics (HEP) similar behavior has now been observed. To aid further study, we present a systematic procedure to derive robust scaling laws and compare design choices on the relevant budget axes for HEP tasks. We first validate the full scaling trajectory on toy problems and then apply the procedure to multi-task transformers on the ~11 billion-jet ATLAS JetSet2 dataset, in both the compute- and data-constrained regimes. For the latter, we predict, to the best of our knowledge for the first time, the jointly optimal model size, training horizon, learning rate and batch size under early stopping. At compute-optimal scaling, we recover a near-equal $\sqrt{C}$ dependence of model and dataset size, and find that auxiliary objectives lower the primary jet-classification loss at equal compute budget. Expanding the inputs toward lower-level data systematically lowers the loss while leaving the scaling exponent nearly unchanged. The onset of the power-law regime is itself set by scale: below a threshold in dataset size the loss carries little information about high-compute scaling, underscoring the value of large, high-quality full-simulation datasets as a foundation for scaling studies and the development of foundation models in HEP.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Vigl, M., Pond, N., Barr, J., Froch, A., Guest, D., Hartman, N., Kagan, M., & Heinrich, L. (2026). How to scale your HEP ML models: A recipe for robust architecture comparisons at scale. https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale

MLA 9

Vigl, Matthias, et al. "How to scale your HEP ML models: A recipe for robust architecture comparisons at scale." https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale.

Chicago (author–date)

Vigl, Matthias, Nikita Pond, Jackson Barr, Alexander Froch, Dan Guest, Nicole Hartman, Michael Kagan, and Lukas Heinrich. 2026. "How to scale your HEP ML models: A recipe for robust architecture comparisons at scale." https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale.

Harvard

Vigl, M., Pond, N., Barr, J., Froch, A., Guest, D., Hartman, N., Kagan, M. and Heinrich, L. (2026) 'How to scale your HEP ML models: A recipe for robust architecture comparisons at scale', Available at: https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale.

Vancouver

Vigl M, Pond N, Barr J, Froch A, Guest D, Hartman N, et al. How to scale your HEP ML models: A recipe for robust architecture comparisons at scale. https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale

IEEE

M. Vigl, N. Pond, J. Barr, A. Froch, D. Guest, N. Hartman, M. Kagan, and L. Heinrich, "How to scale your HEP ML models: A recipe for robust architecture comparisons at scale," https://omanscience.com/en/articles/how-to-scale-your-hep-ml-models-a-recipe-for-robust-architecture-comparisons-at-scale.