Value Alignment

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

14 new papers

A weekly snapshot of new work published in Value Alignment.

Period ending 2026-09-14

16 new papers

A weekly snapshot of new work published in Value Alignment.

Period ending 2026-09-07

19 new papers

A weekly snapshot of new work published in Value Alignment.

Inside this field

Focused directions

581 papers

Latest in Value Alignment

  1. Overtrained, Not Misaligned

    May 12, 2026Joel Schreiber, Ariel GoldsteinMisalignmentHarmful Fine-Tuning

  2. FedOUI: OUI-Guided Client Weighting for Federated Aggregation

    May 12, 2026Alberto Fernández-Hernández, Jose I. Mestre, Cristian Pérez-Corral +3Federated LearningFedavg

  3. Conformity Generates Collective Misalignment in AI Agents Societies

    May 11, 2026Giordano De Marzo, Alessandro Bellina, Claudio Castellano +2Artificial Intelligence AlignmentConformity

  4. Task-Aware Calibration: Provably Optimal Decoding in LLMs

    May 11, 2026Tim Tomov, Dominik Fuchsgruber, Rajeev Verma +1Misalignment

  5. How Value Induction Reshapes LLM Behaviour

    May 8, 2026Arnav Arora, Natalie Schluter, Katherine Metcalf +1Human ValuesLarge Language Model Safety

  6. Theoretical Limits of Language Model Alignment

    May 8, 2026Lucas Monteiro Paes, Natalie Mackraz, Barry-John Theobald +1Large Language Model AlignmentBregman Divergences

  7. SODE: Analyzing Social Dynamics in LLM Agents

    May 6, 2026Inseo Jung, Yoonseok Oh, Kyungryul Back +2Behavioral AlignmentSocial Reasoning

  8. AI Alignment via Incentives and Correction

    May 2, 2026Rohit Agarwal, Joshua Lin, Mark Braverman +1Artificial Intelligence AlignmentIncentives