Correct

Momentum

10 papers in the last four weeks, up 150% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 32

All topics
CardsList
  1. Right Answer, Wrong Mechanism: Detecting Pernicious Divergence in Causal Interventions

    Sep 30, 2026Beiming Liu, Minjie ChenCausal InterventionDivergence

  2. HARISSA: Inference-Time Self-Checks for Efficient and Safe Local Language Model Deployment

    Sep 29, 2026Kenan Alkiek, Moontae Lee, David Jurgens +1Inference-TimePrefill

  3. Verifier Errors in RLVR: Reward Hacking, Limits of Feedback, and Selective Control

    Sep 28, 2026Christian Moya, Elliott Thornley, Guang LinReinforcement Learning With Verifiable RewardVerifiable Rewards

  4. Correct then Forecast: Observer State-Space Models for Time Series Forecasting

    Sep 27, 2026Alexis-Raja Brachet, Guillaume Clavier--Frémond, Abdelhakim Ziani +2State Space ModelsTime Series

  5. A Wrong Turn Does Not Ruin the Journey: Deviation-Guided Skill Self-Evolution for LLM Agents

    Sep 24, 2026Yichun Feng, Jiawei Wang, Haozhe SunSkill EvolutionLarge Language Model Agents

  6. A Chosen Future Can Still Be Rewritten: Causal Writability in Video Models

    Sep 14, 2026Xingyun Wang, Haomin Zheng, Man Yuan +2Correct

  7. When Models Defer to Wrong Answers: A Robustness Audit of Source-Attributed Cues in Multiple-Choice QA

    Sep 8, 2026Manikandan Ravikiran, Siddharth VohraMultiple-Choice QuestionsCorrect

  8. Right Frame, Wrong Rule: Cultural Cues Expose the Financial Knowledge Gap They Were Meant to Close

    Sep 1, 2026Rania Elbadry, Ahmed Heakl, Saeed Almheiri +12Social NormsInterpretive Layer

  9. The Answer Is Not the Argument

    Aug 31, 2026Will Yeadon, Sergio Juárez, Paul Mackay +5AnswerReasoning Traces

  10. Correct Is Not Governed: Provenance Integrity in Agentic Workflows

    Aug 13, 2026Jesus SalasAgentic WorkflowsAccountability

  11. Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks

    Aug 10, 2026Mingyu Luo, Ming Deng, Zilang Qiu +8Jailbreak AttacksAttack-Success Rate

  12. Every Wrong Answer Counts: Option-Level Psychometrics for LLM Multiple-Choice Benchmarks

    Aug 3, 2026Xiao Fei, Yang Zhang, Sarah Almeida Carneiro +1Psychometric PropertiesMultiple-Choice Questions

  13. Detectors Learn the Wrong Thing: Shortcut-Resistant Adversarial Training Against Physically Realizable Attacks

    Jul 23, 2026Yuanhao Huang, Yilong Ren, Jinlei Wang +3Unsupervised DetectionAdversarial Training

  14. Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts

    Jul 10, 2026Renuka Oladri, Mohan Vamsi Varadaraju Priya, Jerry WuLarge Language Model QuantizationLarge Language Models Fail

  15. Position: Correct Answer, Wrong Mechanism -- When AI Scientists Defend General Claims Their Own Data Contradicts

    Jun 22, 2026Steven Young EuligArtificial Intelligence ScientistsCorrect

  16. When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG

    Jun 22, 2026Sahib JulkaLock-InCorrect

  17. The Wrong Kind of Right: Quantifying and Localizing Misfired Alignment in LLMs

    Jun 17, 2026Naihao Deng, Yiming Feng, Chimaobi Okite +4Large Language Model AlignmentMisalignment Persona

  18. Correct Looks Better: Pairwise Comparisons Reveal Accuracy Rankings

    Jun 8, 2026Mina Remeli, Moritz HardtSpearman CorrelationCorrect

  19. CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

    May 19, 2026Ahmed Heakl, Abdelrahman M. Shaker, Youssef Mohamed +4Reinforcement Learning With Verifiable RewardMathematical Reasoning Benchmarks

  20. Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

    May 10, 2026Zhenghan Song, Yulong Liu, Cheng Wan +4Code GenerationPhysics Simulation

  21. Perturb and Correct: Post-Hoc Ensembles using Affine Redundancy

    May 2, 2026Eleanor QuintPost-HocPer-Node Correction