Course 01
Transformer
- Monday, 26 October 2026
- 12:15 – 13:45
- Online
We start with the main NLP tasks and the architectures used before Transformers. We then look at why recurrent networks struggle with long dependencies, how attention works, and how it leads to the Transformer architecture.
On the programme
- 01Background on NLP and tasks
- 02Tokenization
- 03Embeddings
- 04Word2vec, RNN, LSTM
- 05Attention mechanism
- 06Transformer architecture