Human Oversight of AI

Momentum

13 papers in the last four weeks, up 117% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 121

All topics
CardsList
  1. Comprehension Audits to Mitigate Risks from Automated AI Research

    Oct 7, 2026Ronald J. Bodkin, Bahrad A. Sokhansanj, Gillian K. HadfieldHuman Oversight of AIAI Safety

  2. Careful Judge: Safe and Efficient Human-AI Collaborative Decision Making

    Oct 6, 2026Chenyu Zhang, Rachel Luo, Boyi Li +3Human-AI Decision MakingHuman-in-the-Loop AI

  3. A Case Study in Assuring AI-Written Software

    Oct 6, 2026Lindsey Ferris, Sierra BonillaAI Agent ReliabilityAI Assurance

  4. Process Constitutions and Process Stewards: Towards the Next Generation of BPM for Agentic Organizations

    Oct 5, 2026Amin Jalali, Majid RafieiAgentic WorkflowsAI Agent Governance

  5. Feedback Without the Wait: Piloting a Generative AI Practice Platform in a Large Maths Class

    Oct 1, 2026Lili Chen, Gavin Buskes, Yuxin Ren +1Educational AssessmentEducational Technology

  6. Subjects, Not Authors: The Authorship Hazard in Agentic Dataspaces

    Sep 24, 2026Seungho Lee, Changbin LeeRuntime Enforcement for AI AgentsAI Agent Governance

  7. The Ethics of Artificial Intelligence in Military Operations

    Sep 22, 2026Nicolas Drapier, Florian Mauberger, Aladine Chetouani +1AI AccountabilityAI Governance

  8. Seeing Is Not Perceiving: When Synthetic Consumers Can and Cannot Pretest Visual Marketing

    Sep 22, 2026Yi-Lin Tsai, Yung-Hsiu, LaiHuman Oversight of AIGenerative AI Evaluation

  9. greCAPTCHA: Assessing Understanding as Evidence of Research Authorship Under Generative AI

    Sep 17, 2026Justin Payan, Bálint Gyevnár, Atoosa Kasirzadeh +1Authorship AttributionAutomated Evaluation

  10. When Does AI Augment Work? A Workflow-Level Framework for Human-Agent Collaboration

    Sep 14, 2026AI Collaboration, Jiaying Wu, Caleb Ziems +19Human Oversight of AIHuman-AI Collaboration

  11. Characterizing Bluesky Content Moderation Service: From Automation of Service to Landscape of Harms

    Sep 12, 2026Pushpdeep Singh, Sayeh Jarollahi, Ayan Majumdar +5Algorithmic AuditingContent Moderation

  12. MOONWALK: Mediating Operations with Intent-Evidence-Action Alignment Across Junior-Supervisor Review Workflows in Animation/VFX Pre-Production

    Sep 9, 2026Shih-Yu Lai, Wen-Fan Wang, Sai Ling +3Human Oversight of AIHuman-AI Collaboration

  13. It Is Not My Code Anymore

    Sep 8, 2026Augusto CamargoAlgorithmic AccountabilityCode Generation Evaluation

  14. From Matching Models to Recruiting Agents: A Systematized Narrative Review of AI Recruitment Systems, Evaluation, and Governance

    Sep 7, 2026Ziyi Zhao, Guanzheng WeiAI GovernanceHuman Oversight of AI

  15. Knowing Is Not Enough: Information Retrievability as a Precondition to Effective LLM Oversight

    Sep 2, 2026Xinyu Fu, Narayan Ramasubbu, Dennis GallettaLLM ReliabilityHuman Oversight of AI

  16. Responsible Integration of AI in Cancer Genomics: Barriers, Risks, and Pathways to Trustworthy Clinical Translation

    Aug 31, 2026Bahar İlgen, Yiannos Tolias, Denise Kühnert +5Cancer GenomicsOncology

  17. AI Agents Push Humans Out of the Loop

    Aug 24, 2026Margaret Mitchell, Avijit Ghosh, Samir PassiAI Agent SafetyHuman-in-the-Loop AI

  18. The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

    Aug 10, 2026Srinivas Telukunta, Georgios Nektarios Lilis, Lucio BaronAI Agent EvaluationAI Agent Governance

  19. Towards Assurance Closure in AI-Native Large-Scale Agile Software Development

    Aug 7, 2026Ricardo BrittoAI AssuranceSoftware Engineering Agents

  20. A Human Audit of OpenAIs AI-Generated Mathematical Proofs

    Aug 3, 2026Mikołaj Sienicki, Krzysztof SienickiLLM AuditingHuman Oversight of AI

  21. Rethinking Artificial Intelligence in Medical Imaging: Assumptions, Reality, and Reframing

    Jul 29, 2026Arman Rahmim, Nourhan Bayasi, Xiaoxiao Li +2Medical ImagingMedical Imaging Foundation Models

  22. Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI: Lessons from Policing

    Jul 28, 2026Sam Relins, Daniel BirksAI Safety EvaluationAI Governance

  23. Separating Capability from Permission: A Governance Framework for Agentic AI Autonomy Levels

    Jul 26, 2026Haining Zheng, Qian Dong, Rodolfo K. Depena +3AI Agent GovernanceHuman Oversight of AI

  24. Public perceptions of AI-driven decision-making in healthcare: A structural equation modeling approach

    Jul 21, 2026Leonie Westerbeek, Ernesto de Leon, Julia C. M. van WeertTrust in AIHuman Oversight of AI

  25. Nonuniformity Principle in Human-AI Coworking

    Jul 17, 2026An Luo, Jie DingAgentic WorkflowsHuman Oversight of AI

  26. When Not to Automate: A Formal Protocol for Human Preservation in AI-Optimized Organizations

    Jul 17, 2026Jose Manuel de la Chica Rodriguez, Jairo Rodriguez Arias, Spyridon ChouliarasAI Risk ManagementBusiness Process Automation

  27. Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

    Jul 14, 2026Anna Gatzioura, Vrettos Moulos, Nina BaranowskaAI-Assisted Decision MakingAlgorithmic Fairness

  28. Introducing Human-Centeredness in AI-Assisted Lexicography

    Jul 13, 2026Antonio San Martin, Catherine TrekkerHuman-Centered AIHuman Oversight of AI

  29. Comparing Socio-technical Design Principles with Guidelines for Human-centered AI

    Jul 11, 2026Thomas HerrmannHuman-Centered AIHuman Oversight of AI

  30. Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough

    Jul 10, 2026Gaia Grosso, Vinicius Mikuni, Lukas HeinrichAI for ScienceHuman Oversight of AI

  31. Steerability via constraints: a substrate for scalable oversight of coding agents

    Jul 2, 2026Thomas WinningerScalable OversightCoding Agents

  32. Scaling Trends for Lie Detector Oversight in Preference Learning

    Jul 2, 2026Oskar J. Hollinsworth, Ann-Kathrin Dombrowski, Sam Adam-Day +2Scalable OversightDeception in Language Models

  33. Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

    Jun 30, 2026Keito InoshitaHuman Oversight of AIEmotion Recognition

  34. LLMography: Transforming Human-AI Conversations into Traceability, Oversight, and Auditability Indicators

    Jun 28, 2026Mohammed BousmahLLM AuditingHuman Oversight of AI

  35. The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

    Jun 23, 2026Eileanor LaRocco, Sarah Tan, Adarsh Subbaswamy +4AI AccountabilityClinical Decision Support

  36. AI Scientists as Engines of Discovery: A Case for Development within Reformed Institutions

    Jun 22, 2026Raul Jimenez, Boris Bolliet, Francisco Villaescusa-Navarro +7AI Agents for Scientific DiscoveryScientific Hypothesis Generation

  37. The Algorithmic-Human Manager: AI, Apps, and Workers in the Indian Gig Economy

    Jun 18, 2026Omir Kumar, Krishnan NarayananAlgorithmic FairnessAI Governance

  38. Agentic AI Enhances Physician Trust in Clinical Decision Making

    Jun 16, 2026Zhiling Yan, Zhe Fang, David J King +10Clinical Decision-MakingTrust in AI

  39. Regulating the Machine Contributor: Governance and Policy Alignment in Open Source

    Jun 12, 2026Jassem Manita, Aziz AmariAI Agent GovernanceAI Governance

  40. Fault Lines: Navigating Ethics and Responsible AI Where National Policy Meets Local Practice in Public Sector Transformation

    Jun 11, 2026Sitong Lyu, Shabnam Taghiyeva, Mohit Kukadia +1AI AccountabilityAI Governance

  41. (Human) Attention Is (Still) All You Need: Human oversight makes AI-assisted social science reliable

    Jun 11, 2026Chen Zhu, Xiaolu Wang, Weilong ZhangAI Agent ReliabilityAI-Assisted Scientific Research