or try:
Stage 1 • Input Embeddings
Tokenization & Static Initial Embedding
Special tokens [CLS] (start of text) and [SEP] (end of text) are added automatically.
Live Animation
1. TOKEN
[MASK]
→
[MASK]
2. TOKEN ID
ID #103
→
ID #103
3. EMBEDDING
Matrix Lookup
→
Matrix Lookup
4. VECTOR ROW
[1 × 768] Matrix
[1 × 768] Matrix
Raw Layer 0 Embedding for token: [MASK] (ID: 103)
Showing first 12 of 768 dimensions
*The default embedding gets the context from the surrounding tokens in the following layers and eventually predict the word.