cs.CLApr 28, 2026

Investigating Linguistic Steering: An Analysis of Adjectival Effects Across Large Language Model Architectures

Authors: Lars Malmqvist

Organizations: Research and Implementation

Abstract

Achieving reliable control of Large Language Models (LLMs) requires a precise, scalable understanding of how they interpret linguistic cues. We introduce a rigorous framework using Shapley values to quantify the steering effect of individual adjectives on model performance, moving beyond anecdotal heuristics to principled attribution. Applying this method to 100 adjectives across a diverse suite of models (including o3, gpt-4o-mini, phi-3, llama-3-70b, and deepseek-r1) on the MMLU benchmark, we uncover several critical findings for AI alignment. First, we find that a small subset of adjectives act as disproportionately powerful "levers," yet their effects are not universal. Cross-model analysis reveals a "family effect": models of a shared lineage exhibit correlated sensitivity profiles, while architecturally distinct models react in a largely uncorrelated manner, challenging the notion of a one-size-fits-all prompting strategy. Second, focused follow-up studies demonstrate that the steering direction of these powerful adjectives is not intrinsic but is highly contingent on their syntactic role and position within the prompt. For larger models like gpt-4o-mini, we provide the first quantitative evidence of strong, non-additive interaction effects where adjectives can synergistically amplify, antagonistically dampen, or even reverse each other's impact. In contrast, smaller models like phi-3 exhibit a more literal and less compositional response. These results suggest that as models scale, their interpretation of prompts becomes more sophisticated but also less predictable, posing a significant challenge for robustly steering model behavior and highlighting the need for compositional and model-specific alignment techniques.

Explore similar work

CardsList
  1. Towards Understanding Steering Strength

    Feb 2, 2026Magamed Taimeskhanov, Samuel Vaiter, Damien GarreauActivation SteeringSteering

  2. Compositional Multilingual and Behavioral Attribute Steering

    Sep 8, 2026Hyun Gu Kang, Daniil Gurgurov, Tanja Baeumel +2Activation SteeringSteering