cs.LGSep 12, 2026

Certifying Model Upgrades with Slice-Wise Non-Regression and Incumbent Fallback

Authors: Shengwei Zhang, Tao Wu, Fei Qian

Abstract

An updated model can improve an aggregate metric while degrading a slice that matters to a downstream user. We study checkpoint selection subject to non-regression tolerances relative to a retained incumbent. The central distinction is between failing to detect harm and certifying non-inferiority: the former can release harmful updates with high probability when evaluation is noisy. We give a reproducible release procedure that separates candidate search from independent, paired evaluation and returns the exact incumbent when certification fails. Applying established intersection-union and Learn-then-Test principles, we state finite-sample guarantees for one frozen candidate, a finite candidate library, and a prespecified testing order. A joint release decision does not require a slice-count Bonferroni penalty, although certification power can still decrease with the number of slices. In bounded-score simulations, a no-detected-harm gate releases a harmful candidate in 99.7% of trials in one 32-slice setting, compared with 2.6% for an exact non-inferiority gate at a 5% target. A constructed two-block family yields larger certified utility than a scalar path under matched candidate counts. Public digits experiments, including a subsequent continuation that improves average aggregate accuracy, return the incumbent in every run because certification is underpowered. These results establish an auditable protocol and its limitations; they do not establish benefits on foundation-model or multilingual translation upgrades.

Explore similar work

CardsList
  1. When Can Old Evaluations Certify a New Model? Label-Efficient Release Decisions under Evaluator Drift

    Sep 26, 2026Joyanta Jyoti Mondal, Mridul Banik, Md. Shifatul Ahsan Apurba +1Sequential Hypothesis TestingLLM Auditing

  2. Certify or Refuse: A Cross-Model Map for Selective Risk Control with Coverage Floors under Covariate Shift

    Aug 11, 2026Jiamiao Liu, Dewen Qiao, Yu Zhang +1Selective PredictionCovariate Shift

  3. Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification

    Jul 21, 2026Sen Yang, Yuen-Hei YeungMachine UnlearningMachine Unlearning Evaluation