Synthetic Benchmark

Momentum

1 paper in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 6Week of Sep 21

Latest papers 36

All topics
CardsList
  1. TRACE: Tackling Real-World Resource Assignment Problems via Agentic Heuristic Design

    Oct 1, 2026Jose A. Ayala-Romero, Andres Garcia-Saavedra, Xavier Costa-PerezAutomated Heuristic DesignRepair-Assignment Problem

  2. VideoPhysEdit: Physical Counterfactual Video Editing via Rigid-Body Physical Scene Reconstruction

    Sep 28, 2026Conghan Yue, Yuanjie Chen, Yue Han +4Video EditingScene Understanding

  3. SymbolicArena: A Unified Infrastructure for Benchmark Distillation and Dynamic Evaluation in Symbolic Regression

    Sep 28, 2026Ziwen Zhang, Xiju Wu, Yuheng Jing +9Symbolic RegressionSynthetic Benchmark

  4. Test-Time Scaling via Budgeted Multi-Attribute Verification

    Sep 28, 2026Bo Xue, Ji Cheng, Shen-Huan Lyu +2Token Budget AllocationLarge Language Model Responses

  5. CausalArena: Benchmarking Causal Discovery in the Foundation Model Era

    Sep 10, 2026Zi-Rong Li, Si-Yang Liu, Tian-Zuo Wang +1Causal Discovery MethodsCausal

  6. Improving the Realism of Synthetic Clinical Benchmarks Under Utility Constraints

    Aug 6, 2026Omid Bazgir, Md Nasir, Jacob Hoffman +6Synthetic BenchmarkSimulated Patients

  7. Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

    Jul 30, 2026Zhan'ao Yao, Liang Yin, Zhihao Gao +9Scientific DiscoverySynthetic Benchmark

  8. Inverse Learning of Latent Risk-Neutral Densities from Irregular Option Quotes

    Jul 29, 2026Lennon J. Shikhman, Michael Galarnyk, Aadi Dash +1Mathematical FinanceInductive Bias

  9. DoTime: A Synthetic Benchmark Generator for Interventional and Counterfactual Time Series

    Jul 29, 2026Dennis Thumm, Billy Tim Anthony, Ying ChenCausal InferencesSynthetic Benchmark

  10. Scaling Time Series Classification via XAI-Driven Data Reduction

    Jul 17, 2026Davide Italo Serramazza, Thach Le Nguyen, Georgiana IfrimTime-Series ClassificationTime Series

  11. Operator-Informed Gaussian Processes for Complex Helmholtz Wavefields: From Synthetic Benchmarks to In Vivo Brain Elastography

    Jul 15, 2026Boyuan Deng, Kshitiz Upadhyay, Michael ShieldsComplex WavefieldBayesian Inverse Problems

  12. MMAO-Dyn: A Metabolic Multi-Agent Optimizer for Dynamic Optimization

    Jul 1, 2026Jinliang Xu, Liping MaMetabolicSynthetic Benchmark

  13. Evolutional Math: Cross-Validated Island-Model Genetic Programming for Interpretable Symbolic Regression on Small, Wide Datasets

    Jun 20, 2026Artem AndrianovSymbolic RegressionGenetic Algorithms

  14. Domain-Validity-Gated Metamorphic Testing of Scientific ML Surrogates

    Jun 16, 2026Meng Li, Xiaohua Yang, Jie Liu +1Machine Learning SurrogatesValidation

  15. Qualified Educational Capacity Planning under Heterogeneous Student Support Needs: A Synthetic Benchmark and Decision-Support Framework

    Jun 15, 2026Carlos Eduardo Sanoja, Oscar Enrique Moreno MayzCapacityCompetence

  16. HawkesNest: A Multi-Axis Synthetic Benchmark for Spatiotemporal Pattern Complexity

    Jun 15, 2026Yahya Aalaila, Sumantrak Mukherjee, Gerrit Großmann +1Determinantal Point ProcessSpatiotemporal

  17. CODA-BENCH: Can Code Agents Handle Data-Intensive Tasks?

    Jun 13, 2026Yuxin Zhang, Ju Fan, Meihao Fan +2Data Science AgentsSynthetic Benchmark

  18. Measuring What Matters: Synthetic Benchmarks for Concept Bottleneck Models

    Jun 3, 2026Julian Skirzynski, Harry Cheon, Shreyas Kadekodi +2Concept Bottleneck ModelsSynthetic Benchmark

  19. FinStressTS: A Parametric Synthetic Benchmark for Time-Series Forecasting in Finance

    Jun 2, 2026Jiaze Sun, Kelvin J. L. Koa, Ruiyang Ni +3Time Series ForecastingFinance Benchmarks

  20. Test Time Training for Supervised Causal Learning

    May 28, 2026Zizhen Deng, Jiaru Zhang, Rui Ding +5Causal Discovery MethodsTest-Time Training

  21. A Controlled Synthetic Benchmark for Educational Aspect-Based Sentiment Analysis

    May 25, 2026Yehudit Aperstein, Alexander ApartsinSentimentPedagogical Frameworks

  22. SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations

    May 21, 2026Shuaiqi Wang, Aadyaa Maddi, Zinan Lin +1Synthetic DataSynthetic Benchmark

  23. Divergence-Suppressing Couplings for Rectified Flow

    May 18, 2026Yimeng Min, Carla P. GomesRectified FlowEntanglement

  24. Target-Aware Data Augmentation for SAT Prediction

    May 7, 2026Eshed Gal, Uri Ascher, Eldad HaberSatisfiabilityNp-Hard

  25. Joint Treatment Effect Estimation from Incomplete Healthcare Data: Temporal Causal Normalizing Flows with LLM-driven Evolutionary MNAR Imputation

    May 6, 2026Olivia Jullian Parra, Sara Zoccheddu, David Catalan Cerezo +7Heterogeneous Treatment EffectsCausal Inferences

  26. Exploring Spatial Intelligence from a Generative Perspective

    Apr 22, 2026Muzhi Zhu, Shunyao Jiang, Huanyi Zheng +9Stable Spatial UnderstandingMultimodal Large Language Models

  27. SynthSAEBench: Evaluating Sparse Autoencoders on Scalable Realistic Synthetic Data

    Feb 16, 2026David Chanin, Adrià Garriga-AlonsoSparse Autoencoder FeaturesImproving Sparse Autoencoders

  28. Defining Operational Conditions for Safety-Critical AI-Based Systems from Data

    Jan 29, 2026Johann Maximilian Christensen, Elena Hoemann, Frank Köster +1Safety-Critical ScenariosAi-Based

  29. SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models

    Jan 16, 2026Turhan Can Kargin, Wojciech Jasiński, Adam Pardyl +2Spatial ReasoningStable Spatial Understanding