Deception

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

9 new papers

A weekly snapshot of new work published in Deception.

Period ending 2026-09-14

6 new papers

A weekly snapshot of new work published in Deception.

Period ending 2026-09-07

10 new papers

A weekly snapshot of new work published in Deception.

Inside this field

Focused directions

235 papers

Latest in Deception

  1. Janus: A Benchmark for Goal-Conditioned Information Distortion in LLMs

    Jun 9, 2026Polydoros Giannouris, Mohsinul Kabir, Sophia AnaniadouDeception

  2. Building Comparative Motivation Profiles with Instrumental Interventions

    Jun 6, 2026David Vella Zarb, Rustem Turtayev, Taywon Min +2DeceptionSycophancy

  3. Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment

    May 31, 2026Hamidreza Hasani Balyani, Seyed Pouyan Mousavi Davoudi, Alireza Amiri-Margavi +2HonestyEquilibrium

  4. Position: Every Ground Truth is a Human Construction, not an Objective Truth

    May 28, 2026Charlotte Högberg, Ericka Johnson, Kiri L. WagstaffGround TruthTruth

  5. CuriosAI Submission to the CASTLE Challenge at EgoVis 2026

    May 27, 2026Yuto Kanda, Hayato Tanoue, Takayuki HoriCvprLiar

  6. Proper Scoring Rules for Agentic Uncertainty Quantification

    May 23, 2026Suresh Raghu, Satwik Pandey, Shashwat PandeyScoringTruth

  7. PocketAgents: A Manifest-Driven Library of Autonomous Defense Agents

    May 20, 2026Sidnei Barbieri, Ágney Lopes Roth Ferraz, Lourenço Alves Pereira JúniorDefense StrategiesDeception

  8. Selective Safety Steering via Value-Filtered Decoding

    May 14, 2026Bat-Sheva Einbinder, Hen Davidov, Yee Whye Teh +2Large Language Model SafetySafety