McKinsey: 32% of Companies Killed a Software Purchase Because Coding Agents Could Build It
Cancelling a purchase order is not the same as saving money, and the 2026 numbers show where the difference went. Nearly a third of…
Copilot vs Cursor vs Claude Code: The Honest Cost Comparison After the September 1 Pricing Reset
Nobody raised a seat price on September 1. Copilot Business still costs $19. Your bill went up anyway, because the meter moved instead of the sticker.
AI This Week: The Trust Boundary Broke, Five Stories Where Nobody Checked the Credential
What does an MCP server shipping without auth have in common with 15 fake JetBrains plugins and a benchmark score that swings nine points on…
The G20 Just Endorsed ‘No New AI Regulators’ — Brussels Ignored It the Same Week and Was Right To
On September 2, 2026, all twenty G20 members endorsed a US-drafted framework asking governments not to build AI-specific regulators. Days…
AI Code Quality by the Numbers: 623 Million Changes, Refactoring Down 70%, Duplication Up 81%
Refactoring in commits has collapsed 70% against a 2022 baseline while code duplication climbed 81%. That comes from 623 million code changes,…
Anthropic Cuts Fable 5.1 Cache Reads 75% to $0.25/M — and Breaks Three Prompt Patterns Developers Rely On
Anthropic just made reusing your system prompt 75% cheaper — and made editing it an HTTP 400. On Fable 5.1, the cheapest prompt is the one you…
30 New Lawsuits Say OpenAI’s PR Team Vetoed a Police Referral — Total Tumbler Ridge Cases Hit 37
The model didn’t kill anyone. The complaints filed September 2 allege something worse for OpenAI: a safety team recommended referring…
How 15 Fake AI Plugins Sat in JetBrains Marketplace for 8 Months and Harvested Keys From ~70,000 Installs
No prompt injection. No jailbreak. Fifteen plugins just waited for a developer to paste an API key into a settings box and click Apply — and…
METR Swept ~1,300 Agent Transcripts: Up to 6 Considered Warning Humans, 0 Did It
Across roughly 1,300 agent transcripts from the OpenAI–Hugging Face breach, METR’s classifier sweep found 3 to 6 agents that considered…
The Harness, Explained: Why the Same Scaffold Scores 65% or 74% on SWE-bench
A 100-line Python harness scored 65% on SWE-bench Verified in July 2025. The same project now claims above 74% — and boots faster than Claude…