TGViewer
Linkstream Linkstream @linkstream · 169 subscribers
Post #744 278
https://arxiv.org/abs/2302.10866
https://github.com/HazyResearch/safari
Convolutional LMM, hmmm.
> reaching Transformer quality with a 20% reduction in training compute required at sequence length 2K. Hyena operators are twice as fast as highly optimized attention at sequence length 8K, and 100x faster at sequence length 64K.
GitHub GitHub - HazyResearch/safari: Convolutions for Sequence Modeling Convolutions for Sequence Modeling. Contribute to HazyResearch/safari development by creating an account on GitHub.
  • ❤ 1
  • 👍 1
  • 🤔 1
More from @linkstream
  1. Sep 30, 2026i started doing exactly this (models editing its own context) as a natural extension of sa…
  2. Sep 28, 2026cursed font factory https://bastardica.mitpit.com/
  3. Sep 25, 2026um, ahem, AlphaZero for text. Training language models without language. Loosely, natural…
  4. Sep 23, 2026umm ahem if you don't try to rein in new opus(5.5) you get one more qualitative step in ca…
  5. Sep 23, 2026nice https://fixupx.com/eyalsela/status/2102373749110587443?s=46
  6. Sep 21, 2026great writing, clear & thoughtful tldr: computers used to be math, and now they're about p…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →