cs.AIOct 8, 2026

Writing for the Reviewer: Defensive Writing in GPT Models

Authors: Junchi Liao

Abstract

Researchers increasingly use ChatGPT to revise their papers, and recent GPT versions often narrow or even retract the authors' claims. We call such changes defensive writing when the given material does not support them, and we test two explanations: the model corrects the authors' overclaiming, or it writes for an anticipated reviewer. We ask GPT versions and models from other developers to rewrite paragraphs from papers written before ChatGPT, or to write from an evidence sheet that lists a paper's method and results. Defensive writing grows with GPT version. GPT-6-astra retracts the authors' claims outright, and when it writes from the evidence sheet, it still adds the most ungrounded qualifications. The results favor the anticipated-review explanation, and correcting overclaiming explains only a small part. When the models are only asked to polish, defense stays near the level of the originals; mentioning review raises it, and one round of self-review raises it further. At the same time, fewer than one in ten of the claims GPT-6-astra retracts are overstated. AI reviewers score defensive rewrites higher, while human readers find them harder to read and the authors less certain. Combining AI writing with AI review may amplify this style.

Figures & tables

Appendix figures & tables10 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisions

    Jun 11, 2026Xu Yang, Zhizhou Sha, Junbo Li +10Adversarial AttacksAutomated Peer Review

  2. Beating the Style Detector: Three Hours of Agentic Research on the AI-Text Arms Race

    May 4, 2026Andreas Maier, Moritz Zaiss, Siming BayerAI-Generated Text DetectionLanguage Model Generation Evaluation

  3. Gaming AI-Assisted Peer Reviews Poses New Risks to the Scientific Community

    Jun 8, 2026Lin Li, Qi Zhang, Xander Davies +2Adversarial AttacksAdversarial Attacks on LLMs