cs.LGOct 5, 2026

Lossy Compression of PDE Training Inputs: Field Reconstruction Error Does Not Order the Cost to a Trained Operator

Authors: Huy Hoang Le

Organizations: Powermore Ltd., Da Nang, Vietnam

Abstract

Operator-learning benchmarks are stored at full precision and have grown to terabyte scale. Rate-distortion theory says how many bits the stored field needs, while a practitioner needs to know how accurate an operator trained on the compressed data will be. We show that the first does not determine the second, and measure why, compressing the input fields while targets and test inputs stay at full precision. A solution operator attenuates a perturbation of its input. Pushing a compressed field through a surrogate already trained at full precision measures how much of the perturbation that surrogate transmits. The fraction is consistent with the smoothing behaviour of the underlying equation, and it spans more than two orders of magnitude across PDE families. Field reconstruction error is computed before the attenuation and cannot see it. For operators trained with mean squared error it inverts 36 of 104 cost comparisons across datasets, where a probe built from the same forward passes inverts 12. Two families that PDEBench stores with identical initial conditions differ threefold downstream at identical field error. Under the relative-L2 objective of the reference recipe the separation narrows, while the ordering of the family-level median transmission factors is unchanged. After one full-precision training run, the probe evaluates an entire rate curve by forward passes alone. It ranks datasets and rates consistently across the codecs and architectures we test, while its magnitude does not transfer between them.

Figures & tables

Appendix figures & tables25 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. ELADO: Elliptic PDE Assessment Datasets for Operator Learning

    Jun 18, 2026Frank Ehebrecht, Toni Scharle, Martin AtzmuellerLipschitz ContinuityLong-Tailed Distribution

  2. SCOPE: Observation-Conditioned Full-Target Prediction for Sparse PDE Inference

    Sep 29, 2026Ruichen Xu, Siyao Wang, Fang Wan +8Neural Partial Differential Equation SolversPartial Differential Equations

  3. Operator Boosting Produces Pareto-Efficient PDE Surrogates

    Jun 16, 2026Lennon J. ShikhmanNeural OperatorsNeural Partial Differential Equation Solvers