LECTURES A GRATIS GLOBAL SERVICE
⌕ SEARCH GRATIS GLOBAL ↗
LECTURES
Deep Learning for Natural Language: Transformers, Self-Supervised Learning
SOURCE: YOUTUBE · NO TRACKING UNTIL YOU PRESS PLAY · TROUBLE PLAYING? WATCH AT THE SOURCE ↗

Deep Learning for Natural Language: Transformers, Self-Supervised Learning

77 MIN · EN · STATUS: [ STREAMING ]
RATE THIS
MIT · Hands-On Deep Learning Spring 2024 · LECTURE 8

Rama Ramakrishnan continues MIT's 15.773 Hands-On Deep Learning course with a lecture focused on transformer architectures and self-supervised learning for natural language tasks. Building on earlier sessions, Ramakrishnan works through how transformers process sequential text data, the attention mechanisms that let them model relationships between words, and why self-supervised pretraining on large unlabeled corpora has become the dominant strategy for building language models. The seventy-seven minute session is pitched at a practical, hands-on level consistent with the course's applied focus, aimed at students who already have grounding in neural networks and want to understand the mechanics behind modern NLP systems. Expect whiteboard or slide-based explanation of model internals rather than a conceptual survey, tying the architecture back to how these models get trained and used in practice. Part of MIT OpenCourseWare's Spring 2024 offering of 15.773, taught by Rama Ramakrishnan.

At a glance

Lecture facts

Runtime compared with the other 148 Computer Science lectures
Runtime1 h 17 m
Compared with Computer ScienceLonger than 61%
This series

Hands-On Deep Learning Spring 2024

Every lecture in order, sized by its length.

  • Earlier lectures
  • This lecture
  • Still to come
Lecture 8 of 118 h 40 m before this · 13 h 46 m in total

More from this course

10 LECTURES
Introduction to Neural Networks and Deep Learning; Training Deep NNs

Introduction to Neural Networks and Deep Learning; Training Deep NNs

MIT · 57 MIN
Training Deep Neural Networks (cont.); Introduction to Keras/TensorFlow; Application to Tabular Data

Training Deep Neural Networks (cont.); Introduction to Keras/TensorFlow; Application to Tabular Data

MIT · 78 MIN
Deep Learning for Computer Vision: Building Convolutional Neural Networks from Scratch

Deep Learning for Computer Vision: Building Convolutional Neural Networks from Scratch

MIT · 77 MIN
Deep Learning for Computer Vision: Transfer Learning and Fine-Tuning; Intro to HuggingFace

Deep Learning for Computer Vision: Transfer Learning and Fine-Tuning; Intro to HuggingFace

MIT · 76 MIN
Deep Learning for Natural Language: The Basics

Deep Learning for Natural Language: The Basics

MIT · 77 MIN
Deep Learning for Natural Language: Embeddings

Deep Learning for Natural Language: Embeddings

MIT · 78 MIN
Deep Learning for Natural Language: Transformers

Deep Learning for Natural Language: Transformers

MIT · 77 MIN
Generative AI: Large Language Models (LLMs) and Retrieval Augmented Generation (RAG)

Generative AI: Large Language Models (LLMs) and Retrieval Augmented Generation (RAG)

MIT · 75 MIN
Generative AI: Adapting LLMs with Parameter-Efficient Fine-Tuning

Generative AI: Adapting LLMs with Parameter-Efficient Fine-Tuning

MIT · 78 MIN
Generative AI: Text-to-Image Models

Generative AI: Text-to-Image Models

MIT · 76 MIN