Scalable Oversight

Recent momentum

-50%

2 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-14

1 new paper

A weekly snapshot of new work published in Scalable Oversight.

Period ending 2026-09-07

1 new paper

A weekly snapshot of new work published in Scalable Oversight.

25 papers

Latest in Scalable Oversight

  1. AI Agents Push Humans Out of the Loop

    Aug 24, 2026Margaret Mitchell, Avijit Ghosh, Samir PassiArtificial Intelligence AgentsScalable Oversight

  2. Sharding Prevents LLM Oversight Failures and Adversarial Exploitation

    Aug 5, 2026Victor Akinwande, J. Zico Kolter, Aran NayebiJudgesShard

  3. Regulating autonomous and agentic AI

    Jul 23, 2026Chris Reed, Alex Austria, Anmol Bharuka +3Artificial Intelligence GovernanceRegulation

  4. Scaling Trends for Lie Detector Oversight in Preference Learning

    Jul 2, 2026Oskar J. Hollinsworth, Ann-Kathrin Dombrowski, Sam Adam-Day +2DeceptionPreference Learning

  5. Debate Helps Weak Judges Reward Stronger Models

    May 26, 2026Ethan Elasky, Frank Nakasako, Naman GoyalDebateJudges