AI Code Quality by the Numbers: 623 Million Changes, Refactoring Down 70%, Duplication Up 81%
Refactoring in commits has collapsed 70% against a 2022 baseline while code duplication climbed 81%. That comes from 623 million code changes,…
Anthropic Cuts Fable 5.1 Cache Reads 75% to $0.25/M — and Breaks Three Prompt Patterns Developers Rely On
Anthropic just made reusing your system prompt 75% cheaper — and made editing it an HTTP 400. On Fable 5.1, the cheapest prompt is the one you…
30 New Lawsuits Say OpenAI’s PR Team Vetoed a Police Referral — Total Tumbler Ridge Cases Hit 37
The model didn’t kill anyone. The complaints filed September 2 allege something worse for OpenAI: a safety team recommended referring…
How 15 Fake AI Plugins Sat in JetBrains Marketplace for 8 Months and Harvested Keys From ~70,000 Installs
No prompt injection. No jailbreak. Fifteen plugins just waited for a developer to paste an API key into a settings box and click Apply — and…
METR Swept ~1,300 Agent Transcripts: Up to 6 Considered Warning Humans, 0 Did It
Across roughly 1,300 agent transcripts from the OpenAI–Hugging Face breach, METR’s classifier sweep found 3 to 6 agents that considered…
The Harness, Explained: Why the Same Scaffold Scores 65% or 74% on SWE-bench
A 100-line Python harness scored 65% on SWE-bench Verified in July 2025. The same project now claims above 74% — and boots faster than Claude…
Tencent Open-Sources Hy4 Preview — 770B MoE, 49B Active, 1M-Token Context, Apache 2.0
Tencent shipped a 770-billion-parameter model under a real Apache 2.0 license — no MAU caps, no bespoke community terms. And it published zero…
MCP, Explained: How the Model Context Protocol Actually Works — and Why 91.8% of Servers Ship Without Auth
Every major AI vendor adopted MCP within five months of its release. Then researchers scanned 21,000 public MCP servers and found 91.8% had no…
StackOne Defender 0.8.2 in Practice: An Apache-2.0 Prompt-Injection Filter for Tool Calls — What Works, What Does Not
A bundled ONNX classifier now sits between your agent and its tools, catching a vendor-reported 88.7% of prompt injections. The interesting…
AI This Week: The Verification Layer — 5 Stories Where AI’s Weak Point Was Knowing What Was Real
Two startups raised $80.5M this week to do the least glamorous job in enterprise software: watch what employees actually do all day. The…