Hypothesis

Momentum

8 papers in the last four weeks, up 33% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 65

All topics
CardsList
  1. Judgement in the Age of Jev: From Evaluation Scarcity to Evaluation Abundance

    Oct 1, 2026Richard HillJudgementValue

  2. Cheap to Hypothesize, Costly to Verify: The Defense Surface of Agentic Vulnerability Discovery

    Sep 28, 2026Kaikai Zhang, Zihan Zhang, Yuchong Xie +4Model VulnerabilitiesHypothesis

  3. Agentic High-Dimensional Bayesian Optimization with Hypothesis- and Evidence-Guided Search

    Sep 28, 2026Zhixuan Gao, Ke Xue, Rongxi Tan +2Bayesian OptimizationAgentic Optimization

  4. LLMs are not stochastic parrots: Evidence for meaning-mediated abstraction from conlang-like tasks

    Sep 28, 2026Julia Witte Zimmerman, Calla G. Beauregard, Tabia Tanzin Prama +3Large Language Models FailHypothesis

  5. The Linear Representation Hypothesis Needs a Group Action

    Sep 22, 2026Louie Hong Yao, Yuhao Li, Shengchao LiuRepresentation SpaceInterpretability

  6. The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts

    Sep 21, 2026Manjiang Yu, Hongji Li, Zihan Wang +5HypothesisBasin

  7. Pretrained Medical Representations for the Practical Screening of Drug Repositioning Candidates

    Sep 17, 2026Yuhei Fujioka, Daitaro Misawa, Shingo FukumaMedical World ModelElectronic Health Records

  8. HypoEvolve: Genetic Algorithms Enable Multi-Agent LLMs to Discover Scientific Hypotheses

    Sep 14, 2026Jieyuan Liu, Mengzhou Hu, Jefferson Chen +10Hypothesis-Driven Model ExpansionMulti-Agent Evolution

  9. The Interlingua Hypothesis: LLMs Translate via a Latent Task-agnostic Feature Space

    Sep 1, 2026Jacob Brinton, Jannik Brinkmann, Mark Crovella +1Hypothesis

  10. A Human-AI Theorem Connecting Spontaneous and Field-Induced Mechanisms of Collective Behavior in One Dimension

    Aug 31, 2026Weiguo YinHuman-Ai CollaborationSpatial Photonic Ising Machines

  11. Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

    Aug 31, 2026Eddie Conti, Claudio Daka, Álvaro Parafita +3Feature ImportanceExplainable Artificial Intelligence

  12. Perceive to Hypothesize, Verify to Ground: An Agentic Reasoning Framework for Open-World Geo-Localization

    Aug 30, 2026Yutian Jiang, Ruijie Li, Sisuo Lyu +4Cross-View Geo-LocalizationGrounding

  13. Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity

    Aug 7, 2026Zibo Su, Jing Kong, Ruixing Wang +2Optical FlowVideo Restoration

  14. APCReg: Anatomical-Prior-Guided Coarse-to-Fine CBCT--IOS Registration via Multi-View Projection and Reliability-Controlled Residual Correction

    Aug 7, 2026Xincan Zheng, Yaqi Wang, Zhi Li +4Intraoral ScansImage Registration

  15. Factorized Hypothesis Search for Evidence-to-Taxonomy Retrieval

    Aug 6, 2026Linhai Ma, Ethan F. Wei, Xueqing Peng +3TaxonomyRetrievers

  16. Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation

    Aug 4, 2026Shashwat Sourav, Subhadeep Pal, Markus J. Buehler +4HypothesisSynthesis

  17. BayesContact: Uncertain Pose Estimation via Visuo-Tactile Proposals and Simulation-based Inference

    Jul 17, 2026Aditya Kamireddypalli, Matias Mattamala, Joao Moura +3Category-Level Object Pose EstimationTactile

  18. Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery

    Jul 17, 2026Tianyun Zhong, Wangyi Jiang, Wei Wang +15Hypothesis-Driven Model ExpansionHypothesis

  19. The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests

    Jul 13, 2026Edgar DobribanFalse Discovery RateGaussian Primitives

  20. Boosting with List-Decodable Codes

    Jul 7, 2026Addison Prairie, Li-Yang TanBoostingLearnability

  21. The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

    Jun 30, 2026Hongmin LiCategory-Aware Atomic ClaimsScientific Discovery

  22. The FIL Hypothesis: Inductive Biases Help with Kernel Engineering

    Jun 29, 2026Nikolai Rozanov, Subhabrata Dutta, Preslav Nakov +1Inductive BiasFeedback Loop

  23. Unsupervised Causal Abstractions Discovery

    Jun 17, 2026Théo Saulus, Simon Lacoste-Julien, Dhanya SridharCausal Discovery MethodsCausal

  24. The More the Merrier: Combining Properties for ABox Abduction under Repair Semantics in ELbot

    Jun 17, 2026Anselm Haak, Patrick Koopmann, Yasir Mahmood +1Abductive ReasoningEntailment

  25. From Persistence to Survival: Hypothesis Testing, Effect Sizes and Vectorisation for Topological Features

    Jun 10, 2026Juliette Murris, Bernadette Stolz, Karsten BorgwardtPersistent HomologyTopology

  26. Towards Diverse Scientific Hypothesis Search with Large Language Models

    Jun 9, 2026Haorui Wang, Parshin Shojaee, Kazem Meidani +7Hypothesis-Driven Model ExpansionHypothesis

  27. The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment

    Jun 4, 2026Jiachen Zhao, Zhengxuan Wu, Aryaman Arora +3Emergent MisalignmentHypothesis

  28. Read What You Hear: Reference-Free Hypotheses Evaluation with Acoustic Discrepancy

    Jun 3, 2026Zhihan Li, Hankun Wang, Yiwei Guo +3Autoregressive Text-To-SpeechAcoustic

  29. HypothesisMed: Inference-Time Answer Fusion and Structured Hypothesis-Space Reporting for Biomedical Question Answering

    May 31, 2026Md Motaleb Hossen Manik, Ge WangBiomedical TextMedical Visual Question Answering

  30. Testing the Deliteralization Hypothesis in Human and Machine Translation

    May 25, 2026Malik Marmonier, Rachel Bawden, Benoît SagotNeural Machine TranslationMachine Translation

  31. Ontology-constrained multi-LLM scoring of hypothesis support in the predictive processing literature

    May 23, 2026Hamed Nejat, Alexander Maier, Jesse Spencer-Smith +1Predictive CodingHypothesis

  32. Do Language Models Know What Not to Say? Causal Evidence for Statistical Preemption in LLMs

    May 21, 2026Dongxin Guo, Jikun Wu, Siu Ming YiuLarge Language Models FailLinguistics

  33. Value-Gradient Hypothesis of RL for LLMs

    May 20, 2026Arip Asadulaev, Daniil Ognev, Karim Salta +1Large Language Model Reinforcement LearningCritic-Free Reinforcement Learning

  34. On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective

    May 20, 2026Yue Zhang, Zhiyi Dong, Tommaso Cesari +1Reasoning TrajectoryThoughts

  35. Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks

    May 19, 2026Claire Hobbs, R. Thomas McCoyLanguage AcquisitionSyntactic Structure

  36. Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

    May 18, 2026Christiaan G. A. Viviers, Koen de Bruin, Mirre M. Trines +6Research AutomationMedication Leaflet

  37. VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

    May 17, 2026Hezhe Qiao, Hanghang Tong, Ee-Peng Lim +2Model-Based Multi-Agent SystemsTrajectory-Level Credit

  38. Statistical Unlearning of Distributions: A Hypothesis Testing Approach

    May 15, 2026Aaradhya Pandey, Sanjeev KulkarniExact UnlearningDistributions

  39. Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis

    May 14, 2026William Timkey, Brian Dillon, Tal LinzenSurprisalNatural Language

  40. The Alpha Blending Hypothesis: Compositing Shortcut in Deepfake Detection

    May 11, 2026Andrii Yermakov, Jan Cech, Mario Fritz +1Deepfake DetectionUnsupervised Detection

  41. The Cancellation Hypothesis in Critic-Free RL: From Outcome Rewards to Token Credits

    May 9, 2026Tianhao Cheng, Zeyu Huang, Zihan Qiu +5Critic-Free Reinforcement LearningCredit Assignment

  42. Hypothesis generation and updating in large language models

    May 7, 2026Hua-Dong XiongLLM Reasoning StrategiesHypothesis

  43. The Cylindrical Representation Hypothesis for Language Model Steering

    May 3, 2026Lang Gao, Jinghui Zhang, Wei Liu +7Linear Activation SteeringSteering

  44. CHASE: Competing Hypotheses for Ambiguity-Aware Selective Prediction

    May 2, 2026Kartik Jhawar, Yuhao Geng, Atul N. Parikh +1AmbiguityAbstention

  45. Finite-Sample Analysis of Elimination in Active Hypothesis Testing

    May 1, 2026Ziyuan Lin, Hoang Ngoc Nguyen, Jie Xu +1Two-Sample TestingFinite-Sample

  46. A Probabilistic Framework for Hierarchical Goal Recognition

    Apr 24, 2026Chenyuan Zhang, Katherine Ip, Hamid Rezatofighi +2High-Level Subgoal GenerationHierarchical Reinforcement Learning

  47. Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models

    Apr 20, 2026Seyedali Mohammadi, Manas Gaur, Francis FerraroFeasibilityHypothesis

  48. The Umwelt Representation Hypothesis: Rethinking Universality

    Apr 20, 2026Victoria Bosch, Rowan Sommers, Adrien Doerig +1Unified RepresentationNeural Circuits

  49. A phenotype-driven and evidence-governed framework for knowledge graph enrichment and hypotheses discovery in population data

    Apr 18, 2026Adela Bâra, Simona-Vasilica OpreaKnowledge GraphsLocal Causal Structures