Human-Annotated Benchmark

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

11 new papers

A weekly snapshot of new work published in Human-Annotated Benchmark.

Period ending 2026-09-14

38 new papers

A weekly snapshot of new work published in Human-Annotated Benchmark.

Period ending 2026-09-07

15 new papers

A weekly snapshot of new work published in Human-Annotated Benchmark.

Inside this field

Focused directions

233 papers

Latest in Human-Annotated Benchmark

  1. PACE: A Proxy for Agentic Capability Evaluation

    Jul 2, 2026Yueqi Song, Lintang Sutawika, Jiarui Liu +8Agentic BenchmarksAgentic Evaluations

  2. RedactionBench

    Jun 17, 2026Sean Brynjólfsson, Shashvat Jayakrishnan, Esha Sali +2Human-Annotated BenchmarkNamed-Entity Recognition

  3. LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

    Jun 12, 2026Ravidu Suien Rammuni Silva, Ahmad Lotfi, Isibor Kennedy Ihianle +2TutorsHuman-Annotated Benchmark

  4. Prefill Awareness in Large Language Models

    Jun 10, 2026Andy Wang, Parv Mahajan, David Demitri Africa +3Large Language Model SafetyFrontier Models