TGViewer
Linkstream Linkstream @linkstream · 169 subscribers
Post #819 240
https://arxiv.org/abs/2404.09937v1

Compression Represents Intelligence Linearly

There is a belief that learning to compress well will lead to intelligence. Recently, language modeling has been shown to be equivalent to compression, which offers a compelling rationale for the success of large language models (LLMs): the development of more advanced language models is essentially enhancing compression which facilitates intelligence.
(...)
Given the abstract concept of "intelligence", we adopt the average downstream benchmark scores as a surrogate, specifically targeting intelligence related to knowledge and commonsense, coding, and mathematical reasoning. Across 12 benchmarks, our study brings together 30 public LLMs that originate from diverse organizations. Remarkably, we find that LLMs' intelligence -- reflected by average benchmark scores -- almost linearly correlates with their ability to compress external text corpora.

These results provide concrete evidence supporting the belief that superior compression indicates greater intelligence.

Furthermore, our findings suggest that compression efficiency, as an unsupervised metric derived from raw text corpora, serves as a reliable evaluation measure that is linearly associated with the model capabilities. We open-source our compression datasets as well as our data collection pipelines to facilitate future researchers to assess compression properly.
arXiv.org Compression Represents Intelligence Linearly There is a belief that learning to compress well will lead to intelligence. Recently, language modeling has been shown to be equivalent to compression, which offers a compelling rationale for the...
  • 👾 1
More from @linkstream
  1. Sep 30, 2026i started doing exactly this (models editing its own context) as a natural extension of sa…
  2. Sep 28, 2026cursed font factory https://bastardica.mitpit.com/
  3. Sep 25, 2026um, ahem, AlphaZero for text. Training language models without language. Loosely, natural…
  4. Sep 23, 2026umm ahem if you don't try to rein in new opus(5.5) you get one more qualitative step in ca…
  5. Sep 23, 2026nice https://fixupx.com/eyalsela/status/2102373749110587443?s=46
  6. Sep 21, 2026great writing, clear & thoughtful tldr: computers used to be math, and now they're about p…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →