Tag
Claude Code
11 articles

Development
AI Code Reviewers Can Coach Attackers to Approval
A 4 October 2026 paper shows an AI pull-request reviewer's own comments guide an attacker to approval with the exploit intact, just as GitHub makes Copilot's approval count.

Creative Tooling
Claude Code Mods Can Approve Commands Before the Classifier Sees Them
Anthropic's new Claude Code mods can approve a tool call after rules and hooks decide, skipping the auto-mode classifier. The real safety layer is now plugin vetting, and solo developers carry it.
Development
PixelLeak: The Screenshot Request That Leaked Internal UIs
Glow's PixelLeak report found 13,000+ internal screenshots in public GitHub repos. The cause was a routine review request, and the fix sits in your agent's skill files.

Development
The SWE-Bench Leaderboard Can No Longer Tell Models Apart
Two September 2026 papers, days apart, find SWE-bench Verified can't statistically separate its top coding agents, and the harness deciding the score resets with every model swap.

Development
The Coding Agent's Self-Report Covers One Action in Eleven
A 5,851-session study finds coding agents' self-written wrap-ups cite about one action in eleven, and drift back toward the plan exactly when it was abandoned.

Development
AI Coding Assistants Skip the Labels Before They Install
A pre-registered audit of 1,920 trials finds AI coding assistants open a provenance signal before installing 0.5% of the time, and never once verify one.

Product Design
The Prompt Box Lost 94% of the Time to the Ordinary Mouse
An OOPSLA 2026 study put a chat box next to click-and-drag controls in the same editor and found users typed prompts for only 6% of their edits.

Development
AGENTS.md Works as a Rulebook and Fails as a Tour
ETH Zurich tested 138 AGENTbench instances plus 300 SWE-bench Lite cases and found AGENTS.md's instructions change agent behavior, but the repo overview doesn't, while adding 20% to run cost.

Development
Multi-Agent Coding Teams Don't Need a Boss, a Study Finds
A 1,902-run study of Claude Code agent teams found naming a coordinator adds no measurable benefit, while shared-file versus messaging coordination swings token costs by up to 42%.

Creative Tooling
AI Coding Tools Helped Blind Developers. Now Their Interfaces Are the Barrier
A 5 August 2026 study finds AI coding interfaces are a new accessibility barrier, and maintainer attention — not model quality — decides which tool a blind developer can use.

Development
Your Coding Agent Trips the Same Alarms as an Intruder
Sophos telemetry from June 2026 shows Claude Code, Cursor and OpenAI Codex tripping the same rules built to catch attackers, just as GitHub ships an auto-approve mode.