Transformer Under the Hood Live Attention Engine

Attention Map: Softmax(Q·Kᵀ / √d_k)

Tensor [10, 10]
Color Intensity = Attention Weight Hover any cell to inspect mathematical dot-product