Multiple Hypothesis Testing

Latest papers 20

All topics
CardsList
  1. Where Do Two Populations of Persistence Diagrams Differ? Calibrated Local Inference at a Fixed Budget

    Oct 6, 2026Pramita Bagchi, Edward Bae, Atish Mitra +3Two-Sample TestingConfidence Region Estimation

  2. How Sensitive Are LLM Leaderboard Claims to Hidden Model Selection?

    Sep 23, 2026Chen Yang, Xianyang Zhang, Jun ChenLLM EvaluationSelective Inference

  3. Selection-Aware Stress Testing for Interactive Agents

    Aug 31, 2026Yang Xu, Chenang Li, Jiefu Zhang +3AI Agent EvaluationMultiple Hypothesis Testing

  4. ARM: Detector-Agnostic Changepoint Attribution with Finite-Sample Error Control

    Aug 3, 2026Chenchen Peng, Mixia Wu, Qijing Yan +2Change-Point DetectionMultiple Hypothesis Testing

  5. Aggregation of Statistical Evidence under Exchangeability

    Jul 17, 2026Antonin Schrab, Rajen Shah, Arthur Gretton +1Sequential Hypothesis TestingMultiple Hypothesis Testing

  6. The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests

    Jul 13, 2026Edgar DobribanFDR ControlMultiple Hypothesis Testing

  7. Finite Resources False Discovery Rate Control in Structured Hypothesis Spaces

    Jun 13, 2026Binyamin Perets, Shie MannorFDR ControlMultiple Hypothesis Testing

  8. Measurement Under Selection: Decoy-Calibrated Failure Audits for Language Models

    Jun 8, 2026Vyzantinos Repantis, Ameya Gawde, Harshvardhan SinghLLM EvaluationLanguage Model Error Detection

  9. Provable Joint Decontamination for Benchmarking Multiple Large Language Models

    May 20, 2026Zhenlong Liu, Hao Zeng, Hongxin WeiLLM EvaluationBenchmark Contamination

  10. Everywhere Valid Bounds on False Discovery Proportions in Conformal Inference

    May 20, 2026Ziang Song, Ying Jin, Emmanuel J. CandèsFDR ControlConformal Prediction

  11. Controlling False Discovery in Arbitrarily Structured Hypothesis Spaces via Reproducing Kernels

    May 17, 2026Binyamin Perets, Shie MannorFDR ControlMultiple Hypothesis Testing

  12. A Regret Perspective on Online Multiple Testing

    May 13, 2026Qingyang Hao, Kongchang Zhou, Fang Kong +1FDR ControlRegret Minimization

  13. Uncertainty Quantification for LLM-based Code Generation

    May 12, 2026Senrong Xu, Yuhao Tan, Yanke Zhou +6Uncertainty QuantificationLLM Uncertainty Estimation

  14. MinShap: A Modified Shapley Value Approach for Feature Selection

    Apr 16, 2026Chenghui Zheng, Garvesh RaskuttiShapley Value AttributionFeature Selection

  15. Provable FDR Control for Deep Feature Selection: Deep MLPs and Beyond

    Dec 4, 2025Kazuma SawayaFDR ControlFeature Selection

  16. Reliable Selection of Heterogeneous Treatment Effect Estimators

    Nov 23, 2025Jiayi Guo, Zijun GaoCausal Effect EstimationConditional Average Treatment Effect Estimation

  17. Structural Enforcement of Statistical Rigor in AI-Driven Discovery: A Functional Architecture

    Nov 10, 2025Karen SargsyanFDR ControlFormal Verification

  18. Multiple Testing of Linear Forms for Noisy Matrix Completion

    Dec 1, 2023Wanteng Ma, Lilun Du, Dong Xia +1FDR ControlMatrix Completion