Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

YouTube

Transformer Neural Networks, ChatGPT's Foundation, Clearly Explained

StatQuest with Josh Starmer via YouTube

Overview

Google, IBM & Meta Certificates – 40% Off
One Coursera Plus subscription covers most Professional Certificates on Coursera.
Unlock All Certificates
This course explains how transformer neural networks work, from word embeddings and positional encoding through self-attention, encoder-decoder attention, and token decoding. It also explains how transformers support parallel computing.

Syllabus

Awesome song and introduction
Word Embedding
Positional Encoding
Self-Attention
Encoder and Decoder defined
Decoder Word Embedding
Decoder Positional Encoding
Transformers were designed for parallel computing
Decoder Self-Attention
Encoder-Decoder Attention
Decoding numbers into words
Decoding the second token
Extra stuff you can add to a Transformer

Taught by

StatQuest with Josh Starmer

Reviews

Start your review of Transformer Neural Networks, ChatGPT's Foundation, Clearly Explained

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.