RSS Amplifier

Topic · architecture explained

architecture explained

The 15 most recent episodes and tracks on this topic.

Saves to your Watch queue, to pick up on another day or another device.

Pick anything below and it plays in the bar at the foot of the window — and keeps playing while you go on browsing the directory.

  1. L15: Pre-training & fine tuning | gpt architecture adaptation & text generationIntroduction to large language modelsNotes
  2. L14: Causal language modelling | decoder only transformers & autoregressive text generationIntroduction to large language modelsNotes
  3. L13: Transformers for language modelling | encoder only decoder only & encoder decoder architecturesIntroduction to large language modelsNotes
  4. L11: Language modelling | pre training foundation for large language modelsIntroduction to large language modelsNotes
  5. L12: Introduction to language modelling - motivation | fine tuning in GPT & transformersIntroduction to large language modelsNotes
  6. L10: Layer normalization | normalization in transformers encoder decoder architecture explainedIntroduction to large language modelsNotes
  7. L8: Batch normalization | residual connections and layer normalization in transformersIntroduction to large language modelsNotes
  8. L5: Sinusoidal encoding & sequence orderIntroduction to large language modelsNotes
  9. L6: Zooming into decoder layer | decoding transformers masked self attention &cross attentionIntroduction to large language modelsNotes
  10. L7: Positional encoding motivation methods & limitationsIntroduction to large language modelsNotes
  11. L9: Teacher forcing & masked attention | autoregressive decoding with masking in transformersIntroduction to large language modelsNotes
  12. L4: Multi-headed attention in transformers explainedIntroduction to large language modelsNotes
  13. L3: Self-attention in transformers encoder & contextual word embeddingsIntroduction to large language modelsNotes
  14. L2: Attention is all you need transformer architecture explainedIntroduction to large language modelsNotes
  15. L1: Introduction to transformer architectureIntroduction to large language modelsNotes