Category
AI Models
Frontier and open-weight model releases — capabilities, benchmarks, pricing, and what changed.
Inherent Says Its 27B Agent Beat Claude Opus and GPT-5.5
Inherent's Faraday, on a 27B Qwen 3.6, reportedly beat frontier agents at paper replication. Plus Ora's harness benchmark pointing the same way.

GPT-5.6 Sol Is 50% Off Until Sept 18, And o3 Retires Aug 26
Vercel is discounting gpt-5.6-sol by 50% through September 18 via AI Gateway, with no code change. OpenAI retires o3 on August 26. Two clocks, one config.

Open-Source LLM Routers Tested: Three Barely Read Prompts
A common-protocol test of four open-source LLM routers found three emit near-constant tiers, and an always-mid baseline matched one of them exactly.

Needle 2 Packs an Agentic LLM Into a 14MB Binary
Cactus Compute's Needle 2 runs tool-calling in 28MB of RAM — but fine-tuning silently disables the confidence gate that makes it safe to trust.

GLM 5.3 Launches With Big Coding Gains, No API Yet
Z.ai's GLM 5.3 jumps to 28.3 on Terminal Bench 3.0 from GLM 5.2's 4.6, but the standalone API and open weights are still weeks away from release.

Google Lets Users Turn Off Gemini's Visible AI Watermarks
Google is rolling out a toggle to remove visible watermarks from Gemini-generated images, video, and music - but invisible SynthID and C2PA tags stay.

Gemini 3.7 Flash Ships At Half Price, Bigger Coding Gains
Google cut Gemini 3.7 Flash token prices in half and posted double-digit benchmark gains over 3.6 Flash. Here is what changed, and what it costs after 2026.

Anthropic Watermarks Claude Text and Images for EU Rules
Anthropic will embed invisible watermarks and signed provenance metadata in Claude output to meet EU AI Act rules — and some users are already pushing back.

Unreleased Anthropic Model Advances Riemann Hypothesis Bound
An unreleased Anthropic model pushed a Riemann hypothesis bound from 41.6% to 67.2% in 36 hours, then verified the result with the Lean proof assistant.

Meta's Muse Glimmer: A 30B Local Agent Model for Your GPU
Meta shipped a 30B open agent model on Aug 10, sized for a single consumer GPU under Apache 2.0 — specs, framework support, and what to test first.

OpenAI Pauses Astra Model Over Cybersecurity Threshold
OpenAI paused Astra model development after internal tests found it could independently execute cyberattacks, following a related Hugging Face breach.

DeepSeek's Price Hike Signals the End of Cheap AI APIs
DeepSeek plans to raise API prices across the board, ending a year of cuts. Claude Code costs and a $12/month self-hosted Llama setup show what it means.
