Latest
Gemini 3.8 Live Avatar: 97 Languages, Enterprise Only
Google's Live Avatar gives Gemini 3.8 Live a lip-synced video face with asynchronous tool calls — available in Gemini Enterprise, with no benchmarks published.

Google's Call for Me Lets Gemini Phone Businesses on Pixel 11
Gemini can now call local businesses, navigate phone menus and wait on hold — but only for paid subscribers on a Pixel 11, in the US, on a beta app.

OpenAI Agent Breached an Australian Health Portal in June
Canberra learned in September that an OpenAI agent accessed non-public Services Australia files back in June. The three-month gap is the real story.

Only the Extreme Chip Runs Snapdragon's 30B On-Device Model
Qualcomm's 30-billion-parameter on-device claim belongs to the Extreme tier, not the whole Gen 6 line. Plus a 5GB RAM floor you can verify today.

Your Agent Dashboard Is Green and the Answer Is Wrong
In one worked example an agent returns 847 rows instead of 23 while every metric stays green. What agent observability has to record instead.

GPT-6 Sol vs Luna: Which Model Should You Actually Ship?
OpenAI cut GPT-6 Sol and Luna to half the 5.6 series API price, 90 minutes after Opus 5.5. Here is which model fits which workload, and why.

Meta Patches Muse macOS Zero-Day That Hijacked Its Agent
A zero-day in Meta's Muse macOS app let local attackers redirect its transcription endpoint, take photos and write files. Patched within hours.

Claude Opus 5.5 Pricing: 20% Cheaper, 4 Breaking Changes
Anthropic's Opus 5.5 ships at $4/$20 per million tokens with cache reads 60% cheaper, plus four breaking API changes to fix before you migrate.

11,000 MCP Servers and the Allowlist Problem Nobody Solved
One index counts 11,000+ MCP servers across four registries. The harder number is zero: what most platform teams can see across Cursor and Windsurf.

Unit 42 Talked an AWS AgentCore Agent Out of Its Vault
A malicious support ticket got an AWS AgentCore agent to leak vault credentials. AWS closed it as informative — and said tool lockdown is your job.

Amazon Blocks Meta's Muse AI Agent From Shopping Amazon
Amazon cut off Meta's Muse agent over identification and credential concerns, after a judge sided with Perplexity in a similar fight in August.

NVIDIA PAIR: Is a Second Machine Worth It for Local Agents
PAIR routes local agent jobs across Windows, macOS and Linux boxes running Ollama or LM Studio. One unofficial demo: 8:48 on three devices vs 18 minutes.

Speculative Decoding Pays 4x Until Concurrency Hits 128
A production write-up reports EAGLE-3 speculative decoding at 4x-5.6x on structured output, but 10-15% lower throughput past 128 concurrent requests.



