← Back to Labs
LLM Context Windows & Attention Limits
Step through QK^T attention matrix calculation, RoPE positional encoding, KV cache allocation, and needle-in-a-haystack limits
STEP 1 OF 6
Token Vector Projection & RoPE Encoding
Tokens are mapped to dense vector representations, multiplied by Query, Key, and Value weights (Wq, Wk, Wv), and rotated in 2D vector pairs using Rotary Position Embedding (RoPE).
Arrow keys to navigate · R to reset
Tap dots to jump to any step