Who's Accountable When an Agent Gets It Wrong
Courts and regulators have settled this faster than most workplaces have: the organisation that deployed the tool owns what it did. Here's what that means for the person who pressed go.
6 min read10 posts across Supervising Agents, Judging AI, Directing AI
Courts and regulators have settled this faster than most workplaces have: the organisation that deployed the tool owns what it did. Here's what that means for the person who pressed go.
6 min readNoma Labs found that one word placed in a public GitHub issue was enough to trick an AI agent into leaking private repositories. The same week, a fully autonomous AI ransomware operation completed an entire attack chain in 31 seconds. The risk isn't the AI — it's the permissions.
6 min readOpenAI's own safety guidance for ChatGPT Work admits the system can produce 'finished mistakes at a scale that is harder to catch.' The same week, Wharton research killed prompt tricks and Cursor data showed a 46x productivity gap. The differentiator is now verification, not prompting.
6 min readUnit 42 tested 685,339 prompts and found 2.1 million AI-generated URLs — 13,229 already live and malicious. In one case, researchers predicted which domain an AI would hallucinate. Twenty-three days later, an attacker registered it and deployed a phishing kit.
4 min readFigma's CEO Dylan Field told Stratechery that AI output 'draws from the middle of the distribution' — and that teams become viscerally attached to their first concept. The iterate-and-refine skill is what separates competent AI output from distinctive work.
4 min readFord rehired 350 engineers after AI quality inspection produced billions in costs. The same week, a Munich court ruled AI-generated false claims are the deploying organisation's own speech. The gap between 'the AI made an error' and 'you submitted it' is closing.
5 min readAI's most attractive use is work you couldn't do yourself — which is exactly where you can't assess the result. There are real checks available, and they're structural rather than technical.
6 min readA Derbyshire officer was removed from duty and placed under criminal investigation for allegedly using AI to fabricate evidential material — the UK's first criminal case of its kind. The failure mode that caused it appears in professional documents every week.
5 min readTwo South African officials were suspended after 102 of 148 citations in a Cabinet-approved report were AI-fabricated. Nature retracted a 504-citation paper on the same day. The accountability gap between 'AI wrote it' and 'I submitted it' is closing. Here's what you do about it.
4 min readSullivan & Cromwell apologized to a federal judge after a court filing contained 42 AI-generated errors: fabricated citations, misquoted statutes, authorities that don't exist. The opposing firm caught them. Here's what that means for anyone using AI for research, reports, or client-facing work.
5 min readWe use one cookie to keep you signed in — that one isn't optional. Separately, we'd like to record which drills people finish and where they get stuck, so we can fix the parts that don't teach. No ad networks, no selling anything on. Privacy policy.