Transformer Architecture Explained
Under The Hood
0:00 / 0:00
Transformer Architecture Explained
19 960 просмотров · 10 месяцев назад
Under The Hood
17,5 тыс. подписчиков
19 960 просмотров · 10 месяцев назад
Transformer Architecture Explanation from the paper: Attention is all you need.
Watch each components of Transformer Architecture in Detail:
1) Tokenization
• LLM Training Starts Here: Dataset Preparat...
2) Embeddings
• What Are Word Embeddings?
3) Attention Mechanism
• How Attention Mechanism Works in Transform...
Read Original Paper Here:
https://arxiv.org/abs/1706.03762
Timestamp:
0:00 - Introduction
1:15 - Dataset Preparation
2:15 - Encoder: Tokenization, Embedding, PE
5:50 - Encoder: Attention Mechanism
10:05 - Encoder: MHA, Add & Norm, FFNN
13:20 - Decoder: Tokenization, Embedding, PE, MMHA
16:27 - Decoder: Cross Attention, Output
18:05 - Transformer Inference