bahdanau-attention-2014-precursor

IN premiseentries/2026/06/21/wiki-Transformer_deep_learning_architecture-chunk-7.md

Created 2026-06-21T09:55:55+00:00

Bahdanau et al. (2014) introduced additive attention for neural machine translation in RNN encoder-decoder models, predating the Transformer by 3 years