Attention weight (bright = high) Selected query row

Transformer Self-Attention: Query/Key/Value Heatmap (2D)

This 2D companion runs the exact scaled-dot-product self-attention computation behind the 3D scene on a plain canvas built for reading the math rather than orbiting a scene: fixed token embeddings are projected into real query, key and value matrices, scores are computed as Q·Kᵀ/√d_k, and a token-by-token heatmap shows the genuine softmax attention weights — click any token to see exactly which other tokens it attends to and by how much, with every row's weights verified to sum to 1.0.