Attention Explorer (Text-Aware) index462
This page ingests the shared text from index461. The sequence length slider is capped by your current word count. Adjust dk & temperature τ to see how attention sharpens or smooths. The generated random embeddings are stable unless you re-randomize.
Editable Text & Parameters
Words: {{meta.total}}
Unique: {{meta.vocab}}
H(p): {{meta.entropy | number:3}}
PPL: {{meta.perplexity | number:1}}
Parameters
\(\mathrm{Attention}(Q,K,V) = \operatorname{softmax}\left( \frac{QK^{T}}{\sqrt{d_k}\,\tau} \right) V\)
Entropy (text): {{meta.entropy | number:3}}
Perplexity: {{meta.perplexity | number:1}}
|V|: {{meta.vocab}}
Higher text entropy generally implies more diverse keys—making sharp peaks less likely unless dk is large or τ small.
Attention Matrix
Row softmax distributions (L × L). Hover for weight.
q {{hoverInfo.q}} → k {{hoverInfo.k}} : w = {{hoverInfo.w | number:4}}
Value Output (Aggregated)
{{valueVectors | json}}