Catching translation hallucinations with attention misalignment
Output entropy tells you a translation model is unsure but not why. This walks through bidirectional cross-attention, the misalignment features it exposes, and a small QE head that scores tokens without retraining the NMT model.