TGViewer
🚨 AI News | TestingCatalog 🚨 AI News | TestingCatalog @testingcatalog · 7.74K subscribers
Post #9549 763
Mistral Large 4 scores 62% on DeepSWE, outperforming GLM-5.3, according to VentureBeat.

Additionally, it scores 67% on Finch (Financial tasks, SOTA open-weight).

Le Chonk also scores 15% on Harvey’s Legal Agent Benchmark (Legal tasks, SOTA open-weight).

We need a tech report now 👀
  • ❤ 10
  • 👍 2
  • 🗿 1
More from @testingcatalog
  1. Oct 6, 2026Hark Pro is now available on web, iOS, and Android and a $100 Pro plan is now available fo…
  2. Oct 6, 2026GOOGLE 🔥: Nano Banana 2.1 is rolling out on Google AI Studio, APIs, and Gemini! > "High-q…
  3. Oct 6, 2026Mistral launches Large 4 preview with 1 T parameters Mistral Large 4 jumps to the 6th spot…
  4. Oct 6, 2026BREAKING 🔥: Mistral announced Mistral Large 4 "Le Chonk", a new 1T-parameter open-weight…
  5. Oct 6, 2026MISTRAL 🔥: A new big model from Mistral is about to drop, according to Reuters. “The mode…
  6. Oct 6, 2026Google is making Markdown files natively supported in Google Drive and Google Docs. Markdo…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →