TGViewer
🚨 AI News | TestingCatalog 🚨 AI News | TestingCatalog @testingcatalog · 7.63K subscribers
Post #9200 1.38K
OPENAI 🔥: Astra scored 63% on ARC-AGI-3 with a standard harness.

> GPT-6 Astra surpasses the human baseline in action efficiency on ARC-AGI-3. It used fewer actions than the median tested human on 96% of levels.

> A key behavior observed in GPT-6 Astra was its ability to turn unfamiliar environments into compact symbolic world models. It represented game mechanics as logical rules and developed its own domain-specific language shorthand to track state and plan actions.
  • 👍 10
  • ❤ 3
More from @testingcatalog
  1. Sep 27, 2026Meta prepares screen viewing and web voice calls for Muse Meta is preparing two Muse web a…
  2. Sep 27, 2026New Google Flow build now points to Nano Banana 2.1 Google has relabeled references to Nan…
  3. Sep 26, 2026BREAKING 🔥: OpenAI has teased “o” always-on agents to arrive during the DevDay event next…
  4. Sep 26, 2026OpenAI prepares to expand Ultrafast API to more users OpenAI is preparing to expand its Ul…
  5. Sep 26, 2026OpenAI to announce "o" always-on agent during DevDay OpenAI may unveil “O,” an always-on a…
  6. Sep 26, 2026OPENAI 🔥: An upcoming always-on assistant from OpenAI will be named "o". Its reference ap…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →