Language Model Editing

Latest papers 37

All topics
CardsList
  1. From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing

    May 1, 2026Wei Liu, Hongkai Liu, Zhiying Deng +2Knowledge EditingModel Editing

  2. One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them

    Apr 18, 2026Ali Holmov, Paul Youssef, Nandi Schoots +1Knowledge EditingTransformer Interpretability

  3. The GRADIEND Python Package: An End-to-End System for Gradient-Based Feature Learning

    Feb 27, 2026Jonathan Drechsel, Steffen HerboldLLM InterpretabilityLanguage Modeling

  4. Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation

    Nov 8, 2025Yinjie Cheng, Paul Youssef, Christin Seifert +2Knowledge EditingLLM Fine-Tuning

  5. BTBR: A Bayesian-Theory-Driven Probabilistic-Fuzzy Framework for Implicit Bias Removal in Large Language Models

    Aug 20, 2024Yongxin Deng, Xiaoyu Tan, Jing Pan +3Social Bias in Language ModelsLanguage Model Bias Evaluation

  6. LaSEr-Edit: Localized Span-level Error Editing with Energy-based Localization

    Jun 30, 2024Hye Ryung Son, Saehee Eom, Mooho Song +1Controllable Text GenerationLLM Safety