math.OCOct 1, 2026

Reinforcement Learning to Accelerate Primal-Dual Hybrid Gradient for Linear Programming

Authors: Jinhwan Sul, Alex Oshin, Evangelos A. Theodorou

Organizations: Georgia Institute of Technology, USA

Abstract

Primal-dual hybrid gradient (PDHG) methods solve large-scale linear programs (LPs) using GPU-friendly matrix-vector products and projections, but their practical performance depends on coordinating algorithm parameters, acceleration, and restarts. We introduce GALLOP, which uses reinforcement learning to jointly learn continuous algorithm parameters and discrete restart decisions without differentiating through the solver. Its generalized accelerated PDHG update combines separate primal and dual extrapolation, history corrections, and restart anchoring with independently adjustable coefficients. We train a dimension-agnostic feedback policy using a groupwise proximal policy optimization objective that clips likelihood ratios separately for different control groups and excludes inactive acceleration controls on restart transitions. We evaluate GALLOP on six LP families and a public item-placement benchmark. On the main evaluation settings across the six families, GALLOP reduces iteration counts by factors of 1.91.9-5.65.6 and achieves up to a 16.0×16.0\times speedup in algorithm wall-clock time over MPAX. With one policy trained per family, the learned policies generalize without retraining to within-family LPs 3×3\times-400×400\times larger than the largest training instances, including Transport LPs with 10.2410.24 million variables.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Parameter Tuning with Generalization Guarantees for GPU-Accelerated Linear Programming

    Jun 7, 2026Siddharth Prasad, Dravyansh SharmaLinear ProgrammingPrimal-Dual Methods

  2. Learned Preconditioning for a Primal-Dual Interior-Point Method

    Sep 28, 2026Abhinav Madabhushi, Jialin Liu, Minxin ZhangPrimal-Dual MethodsSpectral Preconditioning

  3. Resource-Adaptive Stochastic Gradient Descent for Online Linear Programming without Re-solving

    Sep 23, 2026Jiameng LyuLinear ProgrammingLearning-Augmented Algorithms