Coherence head. On multilingual input, this head keeps the generated output in the target language, preventing the model from slipping into the wrong language partway through a response. It does so by preferentially weighting attention toward same-language tokens. Amber tier: shown conceptually on language-tagged tokens; the within-language weighting is the documented behavior, the specific weights are illustrative.