Multimodal large language models answer audio questions, but how they represent auditory semantics and use them in decisions remains unclear, limiting our understanding of response formation. We study dog barking in Qwen2.5-Omni-7B using Jacobian lens (J-lens) readout and directional interventions. We define the dog direction as a J-lens-derived hidden-state vector associated with dog; adding or removing its component modulates dog-related information. We find this information decodable without dog/bark prompt cues or animal-identification requirements. Directional interventions change response tendencies and some final answers, with effects concentrated in late-layer states immediately before generation across species classification, vocalization classification, and sound description. The dog direction shows no comparable advantage over controls in animal/other classification. These results provide causal-intervention evidence that the dog direction affects output scores in a task-dependent manner, most consistently at L22 and L24 immediately before generation.
Figures & tables
Concept
Dog rank
Non-dog rank
AUC
animal
1.0
58.5
0.786
dog
55.0
53,028.5
0.937
vehicle
695.0
31.0
0.166
bark
964.5
19,653.5
0.872
Table 1: Median full-vocabulary concept ranks at the L24 decision position (smaller ranks indicate higher vocabulary positions). AUC is computed from normalized readout scores.
Task
Token
Δread
95% CI
AUC
Fixed response
dog
1.270
[1.133,1.403]
0.873
Fixed response
bark
0.554
[0.494,0.613]
0.865
Loudness
dog
0.544
[0.475,0.613]
0.822
Loudness
bark
0.249
[0.216,0.281]
0.824
Species
dog
3.659
[3.508,3.800]
0.986
Species
bark
0.881
[0.827,0.934]
0.957
Table 2: Separability of dog-barking and non-dog recordings in selected tasks. Δread is the mean readout-score difference; the fixed response is ready.
Task
Operation
E (L24)
Species
Ablation
0.767†
Species
Injection
1.185†
Vocalization type
Ablation
0.119†
Vocalization type
Injection
0.214†
Sound description
Ablation
2.502†
Sound description
Injection
2.432†
Table 3: Mean effect differences E at L24 decision with norm-matched controls. † : meets the prespecified support criteria. Metric scales differ across tasks.