Category
AI Models
Frontier and open-weight model releases — capabilities, benchmarks, pricing, and what changed.
OpenAI Accuses Moonshot AI of Stealing Model Reasoning
OpenAI says 4,000 users sent 16,000 prompts to extract its models' encrypted reasoning, and links part of the campaign to people tied to Moonshot AI.

Mistral Large 4 Is Out, but Le Chonk's Weights Are Not
Mistral shipped its 1-trillion-parameter Large 4 model on Tuesday, but downloadable weights are weeks out and no public benchmarks exist yet.

An Open 180B Model Tops 10 Hugging Face Boards, Runs on CPU
Darwin-180B-RSI leads 10 of 48 official Hugging Face leaderboards, and a 111 GB 4-bit build of it runs GPU-free at 18-21 tokens per second on CPU.

GPT-6 Astra Ultrafast Is 8x Faster. The Price Isn't Public.
OpenAI's GPT-6 Astra Ultrafast runs on NVIDIA Blackwell and claims up to 8x faster token generation, but the launch announcement names no price at all.

Claude Opus 5.5 vs GPT-6.1 Sol: Effort Flips the Winner
Opus 5.5 wins at max effort, GPT-6.1 Sol at medium, and Sol costs less per task in every published row. List prices, a cost-per-turn model, and the catch.

A 400MB Model Ran a Browser Agent on a 2017 Galaxy Note 8
Qwen3-0.6B scored 10/10 on a live browser task from a 2017 Note 8 and 0/3 without a perception layer — all figures self-reported by the layer's team.

OpenAI's DevDay Slate: GPT-6.1 Sol at One-Fifth the Price
OpenAI launched GPT-6.1 Sol at DevDay, claiming near-Astra coding quality at one-fifth the token price, plus reusable Codex cloud environments.

PrismML's 1-Bit LLM Hits Snapdragon Glasses, None Shipping Yet
Qualcomm showed PrismML's 1-bit Bonsai running on Snapdragon AR1 Gen 1 glasses: a 2B vision model, 4x smaller. No glasses have been announced yet.

Gemini 3.8 Live Avatar: 97 Languages, Enterprise Only
Google's Live Avatar gives Gemini 3.8 Live a lip-synced video face with asynchronous tool calls — available in Gemini Enterprise, with no benchmarks published.

GPT-6 Sol vs Luna: Which Model Should You Actually Ship?
OpenAI cut GPT-6 Sol and Luna to half the 5.6 series API price, 90 minutes after Opus 5.5. Here is which model fits which workload, and why.

Claude Opus 5.5 Pricing: 20% Cheaper, 4 Breaking Changes
Anthropic's Opus 5.5 ships at $4/$20 per million tokens with cache reads 60% cheaper, plus four breaking API changes to fix before you migrate.

Claude Fable 5.1 Benchmarks: 27.9 Points, All Agentic
Fable 5.1 leads all nine reported rows, but the gains bunch in agentic execution while CursorBench barely moves. Input and output pricing did not change.
