cs.LGSep 29, 2026

Jaxolotl: A Unified High-Performance Benchmark Suite for LTL-Based Multi-Task RL

Authors: Mathias Jackermeier, Jacques Cloete, Alessandro Abate

Organizations: Department of Computer Science, University of Oxford · Oxford Robotics Institute, University of Oxford

Abstract

Training agents to follow arbitrary instructions is an important goal of multi-task reinforcement learning (RL). Linear temporal logic (LTL) provides a precise and structured formalism for specifying instructions to agents, and has been successfully adopted for training generalist multi-task policies. However, differences in implementations, task distributions, and evaluation protocols make existing methods difficult to compare, while high computational costs limit the scale and statistical reliability of experiments. We introduce Jaxolotl, a unified high-performance benchmark suite for multi-task LTL-RL to address these concerns. Jaxolotl provides a modular, end-to-end JAX implementation of six representative algorithms and four environments, together with newly curated task suites and a standardised, statistically robust evaluation protocol. By precompiling symbolic task representations into static arrays, Jaxolotl enables fully JIT-compiled training and evaluation, achieving end-to-end speedups of up to 220×220\times and supporting controlled comparisons at substantially greater experimental scale. We use this framework to systematically evaluate existing approaches, revealing complementary strengths and limitations: general methods capable of non-myopic reasoning struggle as the number of propositions grows, while methods with stronger scaling rely on environment-specific assumptions and suffer from myopia.

Figures & tables

Appendix figures & tables13 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. PlatoLTL: Scaling LTL-Guided Multi-Task RL

    Jan 30, 2026Jacques Cloete, Mathias Jackermeier, Ioannis Havoutis +1Multi-Turn Reinforcement Learning

  2. JaxMARL: Multi-Agent RL Environments and Algorithms in JAX

    Nov 16, 2023Alexander Rutherford, Benjamin Ellis, Matteo Gallici +18Multi-Agent Reinforcement LearningCpu-Gpu Hybrid Designs

  3. Live LTL Progress Tracking: Towards Task-Based Exploration

    Apr 18, 2026Noel Brindise, Cedric Langbort, Melkior OrnikLinear Temporal LogicsTrajectory Generation