Category
Dev Tools
Editors, CLIs, and IDE integrations that change how developers write and ship code.
Argent vs agent-device: Same Fixes, One-Third the Cost
Both React Native agent tools scored 12/12 on the same planted bugs. The gap was $1.27 versus $0.41 per run, and which one needed setup first.

GitHub Ships Copilot Local Sandboxing: Lock Credentials First
Copilot's local sandbox is generally available in the CLI, app and VS Code Agent Host. Git credential access is the first setting worth checking.

Your Developers Are Pasting Secrets Into Claude Code and Cursor
GitGuardian says AI-service secret leaks jumped 81% in 2025, and the worst channel is prompt traffic that never touches a git repository at all.

Cloudflare Routes Worker Errors Straight to Your Coding Agent
Issues for Workers hit open beta September 30. It hands exceptions and the affected Worker version to a coding agent — triage them before you let it patch.

LMCP Puts 190+ Mac App Tools Into Claude Code for Free
A free Mac app exposes Mail, Teams, Slack and browser tools to Claude Code by reading local app caches instead of cloud APIs. The vendor's limits matter.

OpenAI Gives Plugins a Sidebar, a Sign-In and a Marketplace
OpenAI's DevDay turned ChatGPT into a distribution surface: plugin extensions, Sign in with ChatGPT with 16 partners, and a 30-plus app marketplace.

12 Checks Before an MCP Server Touches Production Agents
A new go/no-go rubric scores MCP servers 0/1 across two tiers of six checks, and says stop at anything under 5/6 before agents touch it.

16,000 Supabase Databases Are Exposing Data to the Open Web
UpGuard found roughly 16,000 Supabase-hosted databases leaking names, addresses, phone numbers and passwords, as vibe-coded apps ship without configuration.

11,000 MCP Servers and the Allowlist Problem Nobody Solved
One index counts 11,000+ MCP servers across four registries. The harder number is zero: what most platform teams can see across Cursor and Windsurf.

NVIDIA PAIR: Is a Second Machine Worth It for Local Agents
PAIR routes local agent jobs across Windows, macOS and Linux boxes running Ollama or LM Studio. One unofficial demo: 8:48 on three devices vs 18 minutes.

Bend 2 vs SPARK: Proof-Checked AI Code, 442 Lines vs 40
Bend 2 has your agent write machine-checked proofs. A rebuttal proved the same demo in ~40 lines of SPARK versus 442. Which one fits your team.

Vercel Sandbox Now Runs Terminal-Bench in Firecracker VMs
Harbor evals run on Vercel Sandbox with one microVM per trial, network policy enforced outside the guest, and model swaps down to a single --model flag.
