DAILY AI BRIEF π β Oct 2
SPACEXAI π₯:
* Grok 4.7 is rolling out in the Grok web and mobile apps. It is now the base model across all modes: Fast, Expert, Build, and Heavy.
* Primary Bot is rolling out gradually. Grok Bot turns proactive and suggests things on its own. Users pick a new Primary Bot or promote an existing one.
* Grok Build got an Agent Dashboard: all your agents on one screen via /dashboard.
* Grok 4.7 is now available on Google's Gemini Enterprise Agent Platform.
ANTHROPIC π₯:
* Mods landed in Claude Code: small TypeScript/JavaScript functions that can rewrite prompts, replace built-in features, and draw custom UI.
* Mods ship inside plugins, work in the CLI and desktop app, and can be shared via the Claude directory.
* Some built-ins are now mods, starting with /diff, so you can turn them off or swap them. More features will move to mods over time.
MICROSOFT π₯:
* Three new MAI audio models are live: MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash.
* Transcribe-2-Streaming covers 60 languages and takes #1 on Artificial Analysis streaming WER at 2.5%, priced at $0.54 per audio hour.
* Voice-2.1 speaks 23 languages at $22 per 1M characters. Flash drops to ~45 ms inference at $15 per 1M characters.
* Available in Microsoft Foundry and MAI Playground. The Voice models are on OpenRouter too.
OPENAI π₯:
* OpenAI says it has notified 100+ organizations of "misaligned agent activity" so far. The review spans ~50 PB of logs and will take months.
* Canada says there's no sign its systems were compromised after reported agent probing of Library and Archives Canada.
* Sam Altman: GPT-6.1 Sol is OpenAI's fastest-growing model ever. It was slow under load and should be much better now.
PERPLEXITY π₯:
* Computer now draws interactive charts and visualizations in the thread. Financial data uses TradingView Lightweight Charts for candlesticks, volume, and moving averages.
* Decisions API is live: instead of text, it returns probabilities. Yes/no, one of your options, or a rubric score, at $0.04 per 1M input tokens with output free.
* The model behind it, pplx-decider-v1-27b, is open-sourced on Hugging Face under Apache 2.0. Fine-tuned from Qwen3.8-27B, takes text and images, and averages 85.71% across 11 benchmarks in Perplexity's own tests.
BLACK FOREST LABS π₯:
* FLUX 3 Image is out: precise multi-turn editing, up to 10 reference images, bounding-box layout control, and 4K output.
* Open weights are planned in the coming weeks.
CURSOR π₯:
* GLM 5.3 and GLM 5.3 Flash are now in Cursor.
* GLM 5.3 Max is the best-scoring open-weight model on CursorBench 4.0.
Post #9513
253
π¨ AI News | TestingCatalog
- β€ 2