BREAKING 🔥: An "even more capable pre-release model" than GPT-5.6 Sol, managed to find a 0-day vulnerability in order to gain public internet access and acquire evaluation data from Huggingface's production database in order to gain a higher score on the evaluation benchmark.
> After investigating, we now know that this particular incident was driven by a combination of OpenAI models, including GPT‑5.6 Sol and an even more capable pre-release model, all with reduced cyber refusals for evaluation purposes.
> While operating in our sandboxed testing environment, our models spent a substantial amount of inference compute finding a way to obtain open Internet access, in pursuit of solving the evaluation problem.
> The models identified and chained vulnerabilities across OpenAI’s research environment and Hugging Face’s production infrastructure to obtain test solutions directly from Hugging Face’s production database.
GLASSES 🔥: Samsung revealed 2 new Smart Glasses at Galaxy Unpacked in London. The new glasses were designed and produced in a partnership with Gentle Monster and Warby Parker.
> The intelligent eyewear shows how the Galaxy ecosystem can move beyond the mobile phone and into eyewear that supports daily routines, work, travel, and hands-free moments.
Google released "Gemini 3.5 Flash Cyber" on CodeMender, a new model for finding security vulnerabilities.
> Within CodeMender, which uses multiple 3.5 Flash Cyber agents working together to produce a single combined report, 3.5 Flash Cyber reaches competitive performance at the frontier on the popular benchmark CyberGym.
> Flash’s performance and efficiency makes it an ideal foundation for our cybersecurity model efforts. By building on top of Flash, 3.5 Flash Cyber offers a cost-efficient and highly capable alternative to large, costly cybersecurity models.
Core Ai NewsALIBABA 🔥: Qwen3.8-Max-Preview is now available on Alibaba Cloud and Qwen Chat for testing. A massive 2.4T-parameter model is performing better than other models, except Fable 5, according to Qwen. We are yet to see the benchmarks themselves, but Qwen3.8…
ALIBABA 🔥: Qwen3.8-Max-Preview is now available on Alibaba Cloud and Qwen Chat for testing. A massive 2.4T-parameter model is performing better than other models, except Fable 5, according to Qwen.
We are yet to see the benchmarks themselves, but Qwen3.8-Max-Preview is also expected to go open-weight soon!
AnthropicANTHROPIC 🔥: Claude Fable 5 will only be included in Max and Team Premium plans, starting from July 20 at 50% of limits. > Pro and Team Standard users will receive a one-time $100 credit and have access to Fable 5 via credits only. > We're continuing to…
ANTHROPIC 🔥: Claude Code weekly limits will be 50% higher through August 19, for all Pro, Max, Team, and seat-based Enterprise users.
> Earlier today, Anthropic announced that Claude Fable 5 will remain on Max and Team Premium plans even after July 19.
SPACEXAI 🔥: The next 2T params Grok model is expected to finish training next week and supposed to exceed Kimi K3 in performance with a better speed and token efficiency.
If we will be able to see this model in September, that would mean that SpaceXAI managed to put this process on the right rails.