Skip to content
#

sparse-activations

Here is 1 public repository matching this topic...

Interactive explainer: BDH's attention has no softmax, so it is exactly a Hebbian synaptic memory — one fixed-size matrix written once per token. Run both forms, watch them agree to 1e-16, then break it with one toggle. DataForge 2026, Pathway track.

  • Updated Sep 8, 2026
  • HTML

Add this topic to your repo

To associate your repository with the sparse-activations topic, visit your repo's landing page and select "manage topics."

Learn more