
Product Design
AI Models Beat Humans on the Design Brief, and Their Screens Look Alike
An 8 October 2026 benchmark of 15 AI models found all passed over 90% of brief criteria, yet their UI designs resembled each other more than human work. Brief-based review cannot see that.
Read the articleFeatured

AI Code Reviewers Can Coach Attackers to Approval
A 4 October 2026 paper shows an AI pull-request reviewer's own comments guide an attacker to approval with the exploit intact, just as GitHub makes Copilot's approval count.
Read the article
Developers Trust AI When It's Easy to Check, and Code That Runs Passes
Stack Overflow's 2026 survey shows developers trust AI only when they can easily validate it. Three September studies suggest the easiest check, running the code, misses what AI gets wrong.
Read the articleLatest Articles

Creative Tooling
Claude Code Mods Can Approve Commands Before the Classifier Sees Them
Anthropic's new Claude Code mods can approve a tool call after rules and hooks decide, skipping the auto-mode classifier. The real safety layer is now plugin vetting, and solo developers carry it.
Read the article
Creative Tooling
Figma's Agent Goes on the Meter, and Discarded Drafts Pay the Bill
From 6 October 2026 the Figma agent draws AI credits it can't price in advance and never refunds on undo. That bills hardest for the throwaway drafts it was sold to produce.
Read the article
Development
MCP Error Messages Written for Humans Hurt the Smartest Agents Most
A 28 September 2026 preprint finds MCP servers' developer-facing error steps make capable agents fail more, and that naming a server tool in the step fixes it.
Read the article