Topic · architecture explained
architecture explained
The 15 most recent episodes and tracks on this topic.
Saves to your Watch queue, to pick up on another day or another device.
Pick anything below and it plays in the bar at the foot of the window — and keeps playing while you go on browsing the directory.
- L15: Pre-training & fine tuning | gpt architecture adaptation & text generationIntroduction to large language modelsNotes
- L14: Causal language modelling | decoder only transformers & autoregressive text generationIntroduction to large language modelsNotes
- L13: Transformers for language modelling | encoder only decoder only & encoder decoder architecturesIntroduction to large language modelsNotes
- L11: Language modelling | pre training foundation for large language modelsIntroduction to large language modelsNotes
- L12: Introduction to language modelling - motivation | fine tuning in GPT & transformersIntroduction to large language modelsNotes
- L10: Layer normalization | normalization in transformers encoder decoder architecture explainedIntroduction to large language modelsNotes
- L8: Batch normalization | residual connections and layer normalization in transformersIntroduction to large language modelsNotes
- L5: Sinusoidal encoding & sequence orderIntroduction to large language modelsNotes
- L6: Zooming into decoder layer | decoding transformers masked self attention &cross attentionIntroduction to large language modelsNotes
- L7: Positional encoding motivation methods & limitationsIntroduction to large language modelsNotes
- L9: Teacher forcing & masked attention | autoregressive decoding with masking in transformersIntroduction to large language modelsNotes
- L4: Multi-headed attention in transformers explainedIntroduction to large language modelsNotes
- L3: Self-attention in transformers encoder & contextual word embeddingsIntroduction to large language modelsNotes
- L2: Attention is all you need transformer architecture explainedIntroduction to large language modelsNotes
- L1: Introduction to transformer architectureIntroduction to large language modelsNotes
This playlist:.m3u.plsAll the feeds behind it
