The Confident Wrong Answer
Fluency and accuracy are unrelated in AI output, and human error-detection is tuned almost entirely to fluency. Here are the tells that survive that mismatch, and the ones that don't.
5 min read
Practical AI skills for people whose job isn't AI. Most posts end with a drill, because reading about a skill isn't having one.
Fluency and accuracy are unrelated in AI output, and human error-detection is tuned almost entirely to fluency. Here are the tells that survive that mismatch, and the ones that don't.
5 min read
Ethan Mollick's new 'Co-Existence' frame replaces 'use AI more' with a harder question: on this specific task, right now, is AI actually better than you? The answer determines how much supervision you apply — and whether your own skills are quietly eroding.
4 min read
Uber blew its annual AI budget in months. Simon Willison's $200 subscription runs $2,180 in compute. Anthropic's revenue went 5x in five months. If your team budgeted for AI tools in Q4, those numbers are already wrong — here's how to reset them.
4 min read
Research published last week showed frontier AI models disagree with each other on 67% of real-world fact-checking claims — 77% in legal contexts, 71% in health. A second model isn't verification. It's a second opinion that contradicts the first most of the time.
5 min read
Three studies published in the same week show that how you use AI determines whether it sharpens or erodes your judgment. Turkish students scored better with AI help and worse without it. BCG consultants followed AI errors they would normally have caught. Taipei students who used AI as a tutor gained six to nine extra months of learning.
5 min read
Robinhood is trading stocks, Gemini Spark is running your calendar overnight, and a CVSS 9.3 vulnerability let five lines of text exfiltrate an entire M365 environment. Four questions to answer before you connect any AI agent to a real system.
5 min read
Josh Comeau watched an expert developer triple their output with AI and watched beginners spend three hours prompting a model to solve something they fixed manually in 30 seconds. The difference wasn't the model. It was what each person brought to it.
4 min read
Google's AI Pointer and Gemini Rambler let you point, speak, and get results without typing a word. The syntax of prompting is being absorbed into the interface. The judgment behind it isn't.
4 min read
VS Code silently stamped Copilot attribution on millions of commits this week — even when Copilot wasn't used. The instinct behind that decision is running in every AI tool you use at work.
4 min read
Practical AI skills, sent when there's something genuinely useful to share. No filler, no hype.
We use one cookie to keep you signed in — that one isn't optional. Separately, we'd like to record which drills people finish and where they get stuck, so we can fix the parts that don't teach. No ad networks, no selling anything on. Privacy policy.