
Fable 5.1 lands: same token price, much cheaper cache reads
Anthropic shipped Fable 5.1 to Claude and Claude Code. Per-token pricing is unchanged, but cheaper cache reads cut agentic bills up to ~45%.
Model releases, benchmarks and capabilities. 106 stories, page 2 of 5.

Anthropic shipped Fable 5.1 to Claude and Claude Code. Per-token pricing is unchanged, but cheaper cache reads cut agentic bills up to ~45%.

Hy4 preview: 770B params, 49B active, 1M context, Apache 2.0 — and it undercuts GLM-5.3 by ~40% on price.

Leaked strings point to Grok Bot landing in Grok's mobile apps, plus a hidden nav tab for custom agents — Heavy tier only.

GPT-Image-2 now accepts background=transparent in preview — alpha-channel assets straight from the API, no cutout step.

A 'Hub Mode' string surfaced in Anthropic's code carrying Orbit's old icon — likely a multi-agent dashboard, still unannounced.
Bot Mode ships default-on in Hermes Agent v0.20.3: each bot gets its own model, memory, skills — and can @mention the others.

Alibaba's Qwen passed 3 billion downloads and now accounts for more than half of all open-model downloads worldwide.

Zhipu's 743B GLM-5.3 beat Mythos 5 and GPT-5.6 Sol on CyberGym at 84.5% — then trailed both by 24 points on ExploitBench.

Google's new workhorse model is now default in the Gemini app — 4 points smarter than 3.6, 40% faster than GPT-5.6 Terra.

Google halved Gemini 3.7 Flash to $0.75/M, OpenAI shipped a 14x-faster tier, and DeepSeek hiked prices 4.5x.

Both shipped Aug 12-13. Grok wins knowledge work, DeepSeek wins agents and security — at $0.87 vs $6 per million output tokens.

SpaceXAI's new frontier model matches GPT-5.6 Sol on benchmarks at a fraction of the price. The fine print doubles it.

Codex now ships native builds for Ubuntu, Debian 13 and Fedora — but Computer Use is still missing from the Linux version.

Meta's first no-strings open-weight model in years: 30B, agentic, beats Gemma 4 31B on coding, runs on one consumer GPU.

Muse Glimmer is Apache 2.0, runs offline on a 24GB card, and beats Qwen and Gemma on multi-step tool use.

Meta's new terminal coding agent hits 59.3% on DeepSWE 1.1 — solid, but Claude Opus 5 leads on Meta's own charts.

Alibaba shipped its biggest model yet — 2.4T params, 1M context, frontier-tier coding scores. Open weights land next week.

DeepSeek-V4-Flash-0731 is live in API beta, outscoring V4-Pro-Preview on agentic benchmarks and adding Responses API support.

Altman demoed Astra to US senators — multiple agents grinding on one hard problem for hours, not one model answering fast.

A Homeworld-style space RTS — code, 3D models and all — generated end-to-end by Claude Opus 5 for 4.6M output tokens.

Moonshot open-sourced a 2.8T-parameter model on July 27. Self-hosting it needs 64+ accelerators, so most devs will just pay the API.

Three weeks after its consumer debut, OpenAI's full-duplex voice model is live for Business, Enterprise and Edu seats.

Kimi K3's weights are public: 2.8T params, 104B active, 1M context. Largest open-weight model ever shipped.

OpenAI added shareable links for custom ChatGPT Pets — copy from Settings Personalisation, and anyone can adopt yours. Web only for now.