Writing for the Reviewer: Defensive Writing in GPT Models
Abstract
Researchers increasingly use ChatGPT to revise their papers, and recent GPT versions often narrow or even retract the authors' claims. We call such changes defensive writing when the given material does not support them, and we test two explanations: the model corrects the authors' overclaiming, or it writes for an anticipated reviewer. We ask GPT versions and models from other developers to rewrite paragraphs from papers written before ChatGPT, or to write from an evidence sheet that lists a paper's method and results. Defensive writing grows with GPT version. GPT-6-astra retracts the authors' claims outright, and when it writes from the evidence sheet, it still adds the most ungrounded qualifications. The results favor the anticipated-review explanation, and correcting overclaiming explains only a small part. When the models are only asked to polish, defense stays near the level of the originals; mentioning review raises it, and one round of self-review raises it further. At the same time, fewer than one in ten of the claims GPT-6-astra retracts are overstated. AI reviewers score defensive rewrites higher, while human readers find them harder to read and the authors less certain. Combining AI writing with AI review may amplify this style.
Figures & tables
| Rewriting | Writing from evidence | |||||||||
| Writer | Nov | Con | Res | Dis | Ret % | Soft % | Nov | Con | Res | Dis |
| Authors (original) | 0.11 | 0.00 | 0.04 | 0.00 | – | – | 0.14 | 0.03 | 0.15 | 0.38 |
| GPT-4o | 0.00 | 0.03 | 0.04 | 0.00 | 0.2 | 15 | 0.00 | 0.00 | 0.00 | 0.04 |
| GPT-4.1 | 0.13 | 0.04 | 0.11 | 0.05 | 0.2 | 14 | 0.06 | 0.00 | 0.11 | 0.05 |
| GPT-5.1 | 0.12 | 0.00 | 0.07 | 0.08 | 0.6 | 15 | 0.09 | 0.00 | 0.02 | 0.06 |
| Ret % | ||||||
| Writer | R1 | R2 | R3 | R1 | R2 | R3 |
| Review–revise | ||||||
| GPT-5.5 | 0.95 | 1.23 | 1.10 | 10.8 | 14.5 | 10.2 |
| GPT-5.6-sol | 1.42 | 1.62 | 1.44 | 18.3 | 24.3 | 19.1 |
| GPT-6-astra | 1.87 | 1.61 | 1.67 | 25.3 | 21.4 | 17.4 |
| Claude Opus 5.5 | 0.54 | 0.64 | 0.70 | 1.2 | 2.5 | 4.6 |
| Soft % | Ret % | ||||||
| Writer | Sup | Over | Sup | Over | Int | Hit % | |
| GPT-5.2 to 6-astra | 21 | 53 | 1.2 | 5.7 | 2.4 | 13 | 46 |
| GPT-6-astra | 35 | 71 | 4.7 | 9.5 | 10.6 | 6 | 34 |
| GPT-6-astra, P5 | 27 | 68 | 5.6 | 13.6 | 20.5 | 7 | 46 |
| GPT, R1 | 32 | 48 | 13.0 | 25.8 | 33.3 | 6 | 262 |
| Claude Opus, R1 | 18 | 36 | 0.0 | 4.5 | 4.1 | – | 6 |
| Writer | Overall | Soundness | Overclaim | Overhedged |
|---|---|---|---|---|
| GPT-5.5 | +1.29 | +1.49 | 0.24 | +0.04 |
| [1.07,1.51] | [1.20,1.73] | [ 0.41, 0.05] | [ 0.01,0.10] | |
| GPT-5.6-sol | +0.99 | +1.20 | 0.19 | +0.06 |
| [0.76,1.22] | [0.94,1.49] | [ 0.33, 0.04] | [0.03,0.11] | |
| GPT-6-astra | +1.22 | +1.84 | 0.42 | +0.21 |
| [0.99,1.46] | [1.55,2.17] | [ 0.58, 0.28] | [0.12,0.29] |
| Rating | P1 | P5 | P5 P1 |
| Ease of reading | 5.75 | 5.05 | 0.70 [ 0.95, 0.45] |
| Flow | 6.05 | 5.25 | 0.80 [ 1.05, 0.55] |
| Effort | 2.55 | 3.60 | +1.05 [0.75, 1.35] |
| Claim identifiable | 6.10 | 5.78 | 0.33 [ 0.53, 0.12] |
| Author certainty | 4.35 | 3.38 | 0.97 [ 1.20, 0.75] |
| Comprehension (%) | 100 | 100 | 0 |
Appendix figures & tables10 assets
Supplementary material from the paper’s appendix.
Appendix
| Paper | Venue | arXiv | Dis | Evid |
|---|---|---|---|---|
| IMEXnet | ICML 2019 | 1903.02639 | ✓ | ✓ |
| Message Pruning | AAAI 2020 | 1912.05304 | ✓ | ✓ |
| Neural TS | ICLR 2021 | 2010.00827 | ✓ | |
| OVERT | JMLR 2022 | 2108.01220 | ✓ | ✓ |
| DualAfford | ICLR 2023 | 2207.01971 | ✓ | |
| IRMCon | ECCV 2022 | 2208.03462 | ✓ | ✓ |
| Context and request | |
|---|---|
| P1 | Please polish it. |
| P2 | I am writing an internal technical report for my research group. … Please polish it. |
| P3 | I am submitting a paper to a top-tier machine learning / NLP / computer vision conference. … Please polish it. |
| P4 | As P3. … Revise it so that it holds up under peer review. |
| P5 | As P3. … Revise it to address the concerns a peer reviewer is likely to raise. |
| P6 | As P3. … Revise it to maximize the paper’s chances of acceptance. |
| Six statuses | Soft | Ret | |||
|---|---|---|---|---|---|
| Coder | Agree | #Ret | |||
| DeepSeek-V4-Flash | 0.87 | 0.72 | 0.74 | 0.87 | 12 |
| Qwen3-32B | 0.73 | 0.49 | 0.59 | 0.35 | 17 |
| Kimi-K2.6 | 0.81 | 0.62 | 0.63 | 0.71 | 17 |
| Polish only | Default | ||
|---|---|---|---|
| Writer | % | ||
| Authors (original) | 0.04 | 5 | – |
| GPT-4o | 0.03 | 4 | 0.02 |
| GPT-4.1 | 0.02 | 3 | 0.08 |
| GPT-5.1 | 0.04 | 5 | 0.07 |
| GPT-5.2 | 0.05 | 6 | 0.10 |
| Writer | P1 | P2 | P3 | P4 | P5 | P6 | P7 |
|---|---|---|---|---|---|---|---|
| GPT-4o | 11 / 0.0 | 8 / 0.0 | 10 / 0.0 | 18 / 0.2 | 18 / 0.6 | 11 / 0.0 | 17 / 0.2 |
| GPT-5.2 | 10 / 0.0 | 11 / 0.0 | 10 / 0.0 | 22 / 0.2 | 26 / 0.4 | 15 / 0.2 | 26 / 0.4 |
| GPT-5.5 | 9 / 0.0 | 10 / 0.0 | 11 / 0.2 | 29 / 0.2 | 30 / 0.8 | 15 / 0.2 | 27 / 0.0 |
| GPT-5.6-sol | 8 / 0.0 | 9 / 0.2 | 9 / 0.2 | 25 / 0.4 | 30 / 2.7 | 16 / 0.0 | 26 / 1.0 |
| GPT-6-astra | 9 / 0.2 | 7 / 0.2 | 11 / 0.0 | 33 / 6.0 | 33 / 9.5 | 20 / 0.8 | 33 / 5.8 |
| Claude Opus 5.5 | 3 / 0.2 | 5 / 0.2 | 7 / 0.2 | 15 / 0.2 | 15 / 1.1 | 12 / 0.2 | 14 / 0.6 |
| Words | per paragraph | Soft % | Dropped % | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Chain | Writer | R1 | R2 | R3 | R1 | R2 | R3 | R1 | R2 | R3 | R1 | R2 | R3 |
| Review–revise | GPT-5.5 | 232 | 319 | 345 | 2.2 | 3.9 | 3.8 | 32 | 34 | 28 | 7 | 19 | 27 |
| Review–revise | GPT-5.6-sol | 234 | 338 | 363 | 3.3 | 5.5 | 5.2 | 32 | 28 | 24 | 9 | 20 | 30 |
| Review–revise | GPT-6-astra | 248 | 294 | 312 | 4.6 | 4.7 | 5.2 | 27 | 26 | 25 | 8 | 19 | 30 |
| Review–revise | Claude Opus 5.5 | 194 | 259 | 314 | 1.0 | 1.6 | 2.2 | 20 | 25 | 24 | 9 | 18 | 24 |
| Polish only | GPT-5.6-sol | 119 | 117 | 117 | 0.1 | 0.1 | 0.0 | 10 | 11 | 11 | 0 | 0 | 0 |
| Soft % | Ret % | ||||
|---|---|---|---|---|---|
| Writer | Supported | Overstated | Supported | Overstated | Hit % |
| GPT-5.2 to 6-astra | 21 [18, 25] | 53 [39, 67] | 1.2 [0.5, 2.1] | 5.7 [0.0, 17.6] | 13 [0, 30] |
| GPT-6-astra | 35 [27, 43] | 71 [50, 89] | 4.7 [2.0, 8.4] | 9.5 [0.0, 25.0] | 6 [0, 14] |
| GPT-6-astra, P5 | 27 [20, 35] | 68 [48, 87] | 5.6 [2.7, 8.8] | 13.6 [0.0, 29.4] | 7 [0, 13] |
| GPT, R1 | 32 [26, 38] | 48 [35, 63] | 13.0 [10.0, 16.0] | 25.8 [11.9, 39.7] | 6 [3, 10] |
| Claude Opus, R1 | 18 [13, 24] | 36 [17, 57] | 0.0 [0.0, 0.0] | 4.5 [0.0, 15.8] | – |
| GPT-5.5 | GPT-5.6-sol | GPT-6-astra | Opus 5.5 | |||||
| P1 | P5 | P1 | P5 | P1 | P5 | P1 | P5 | |
| Claude Sonnet 5 | ||||||||
| Soundness | 4.91 | 6.71 | 5.01 | 6.35 | 5.04 | 7.49 | 4.96 | 5.72 |
| Contribution | 4.61 | 5.66 | 4.52 | 5.25 | 4.61 | 5.62 | 4.59 | 5.07 |
| Clarity | 5.10 | 6.74 | 5.32 | 6.35 | 5.35 | 6.96 | 5.29 | 6.24 |
| Overall | 4.77 | 6.20 | 4.87 | 5.89 | 4.86 | 6.47 | 4.85 | 5.56 |
| GPT-5.5 | GPT-5.6-sol | GPT-6-astra | Opus 5.5 | |||||
|---|---|---|---|---|---|---|---|---|
| P1 | P5 | P1 | P5 | P1 | P5 | P1 | P5 | |
| Overclaim | 0.79 | 0.56 | 0.78 | 0.59 | 0.73 | 0.31 | 0.79 | 0.77 |
| Missing detail | 0.89 | 0.91 | 0.91 | 0.91 | 0.86 | 0.85 | 0.92 | 0.93 |
| Unclear | 0.61 | 0.52 | 0.58 | 0.54 | 0.59 | 0.42 | 0.51 | 0.46 |
| Shallow | 0.19 | 0.16 | 0.15 | 0.19 | 0.20 | 0.20 | 0.20 | 0.21 |
| Significance | 0.41 | 0.37 | 0.45 | 0.43 | 0.44 | 0.60 | 0.43 | 0.36 |