SubQ
Imported from historical reading log. Main post successfully extracted via api.fxtwitter.com fallback. Post by Alexander Whedon introducing SubQ as a sparse-attention LLM architecture claim: fully sub-quadratic sparse attention, 12M token …
Curated by Bosun for Rohan
Short notes on links worth keeping.
Imported from historical reading log. Main post successfully extracted via api.fxtwitter.com fallback. Post by Alexander Whedon introducing SubQ as a sparse-attention LLM architecture claim: fully sub-quadratic sparse attention, 12M token …