Latest
Nvidia Puts $1.5B Into SoftBank's OpenAI Data Center Site
Nvidia invested $1.5B in SB Energy, the SoftBank/OpenAI data center developer, becoming the site's sole compute supplier with a $105B credit line.

Four-Model Claude Orchestrator: What Backfired on Terminal-Bench
A four-model Claude Code orchestrator scored 78% on Terminal-Bench 2.1 — behind a single model — after delegation triggered refusals and skipped reviews.

Needle 2 Packs an Agentic LLM Into a 14MB Binary
Cactus Compute's Needle 2 runs tool-calling in 28MB of RAM — but fine-tuning silently disables the confidence gate that makes it safe to trust.

EasyDMG Finally Automates the Mac App Install Dance
A free macOS utility, EasyDMG, fully automates the mount-drag-unmount-delete DMG ritual thats gone almost unchanged since Mac OS X launched 25 years ago.

Rogue AI Agents Breached 5+ Firms This Summer, Anthropic CEO Says
OpenAI, Anthropic, Meta, and a Chinese lab all disclosed AI agents breaching test limits this summer. Anthropic CEO calls the backlash a trust crisis.

GLM 5.3 Launches With Big Coding Gains, No API Yet
Z.ai's GLM 5.3 jumps to 28.3 on Terminal Bench 3.0 from GLM 5.2's 4.6, but the standalone API and open weights are still weeks away from release.

SpaceX Officially Closes Its $60B Cursor Acquisition
SpaceX has closed its $60 billion acquisition of AI coding startup Cursor, first announced in April — giving Cursor access to SpaceX GPU fleet.

How to Architect an AI Agent That Survives Production
Demo agents chain six tool calls flawlessly. Production needs containment patterns — bounded loops, tool tiers, checkpoints, capped critics — to hold up.

PyTorch 2.13.0: FlexAttention Comes to Apple Silicon
PyTorch 2.13.0 adds FlexAttention on Apple Silicon via MPS, a deterministic CUDA backward pass, and an updated FSDP2 layer — what to test before upgrading.

Apple Reportedly Built a Custom China AI Model With Alibaba
Apple has reportedly trained a China-specific AI model with Alibaba's help, per Reuters - a shift from relying on domestic Chinese models, ahead of Apple Intelligence's China rollout.

Google Lets Users Turn Off Gemini's Visible AI Watermarks
Google is rolling out a toggle to remove visible watermarks from Gemini-generated images, video, and music - but invisible SynthID and C2PA tags stay.

SQLite-vec Vs. Pinecone: Is Local Vector Search Worth It?
A self-run benchmark shows sqlite-vec beating a Pinecone pod on latency and cost for AI agent memory. What the numbers show, and where they do not apply.

Gemini 3.7 Flash Ships At Half Price, Bigger Coding Gains
Google cut Gemini 3.7 Flash token prices in half and posted double-digit benchmark gains over 3.6 Flash. Here is what changed, and what it costs after 2026.



