
SIGNAL / NOISE
Buy the Nines
Anthropic shipped Claude Haiku 5.5 on Wednesday and everyone wrote the same sentence: the cheap model got 90% cheaper. True, and boring. The line nobody circled sits four rows down the page. For the first time on a Haiku-class model you can set the effort level, five of them, low to high. That isn't a price cut. That's a dial.
Here's what the dial does, and why thirty years of running money makes me stop on it. You have never been able to buy correctness by the yard. You hired a person, you got one price and one reliability you couldn't actually see, and you prayed. You could not walk in and say "give me my analyst, but four points more accurate, and tell me what the four points cost." Now you can. Point an agent at a job, tell it what "right" means, and it will quote you the price of being right 90% of the time, 95%, 99%. Correctness just became a line item. You buy the nines the job is worth and not one more.
That changes the whole fight. For two years we argued about which model is smartest. Wrong fight. Intelligence is now a commodity that deflates about 79x a year (we did that math in Two-Way Street), three credible open-weight models landed this week to keep it honest, and even Musk's own bot started routing jobs to Anthropic's Claude when the task called for it. The model is a dial tone. The game moved to operations: what do you need to be right about, how right, and what will you pay for the last nine.
But here's the part that bites the people who skip it. You can only buy a number you can measure, and you can only measure it with a ruler the agent doesn't get to hold. Tell the agent to grade its own homework and it optimizes to look right, not to be right. We watched Gemini do exactly that last week, benchmaxxing the test it knew was coming (No-Win Scenario). This week OpenAI dropped 372 math results and the mathematicians can't tell which are new, because the prompts were withheld. A result you can't check isn't a result. It's a rumor with a decimal point.
So the dial is real, and it's the best news an operator has had in years. It just comes with a bill nobody reads until it's structural: agents make work for agents, checkers checking checkers, and that bureaucracy bloats exactly where no one's watching the meter. The fix is the oldest one there is. Own the ruler, then buy the nines.
At COAI today: the full Signal/Noise, with the operator's three-step playbook and why an agent bureaucracy bloats wherever the decider isn't the payer, is live at getcoai.com.
What "right" means for each agent you've turned loose, what the nines actually cost you, and whether the ruler measuring them is one the model never sees.
ONE — A NUMBER THAT SUMMARIZES THE DAY
90%. That's how far Anthropic cut the list price of its cheapest model this week, down to a dime per million words in. The discount isn't the story. The control panel underneath it is: five effort levels, a dial that runs from cheap-and-good-enough to slow-and-sure. For the first time you can buy correctness the way you buy bandwidth. Pick the nines the job is worth, pay for those, and stop. You could never do that with a human, and most of your work doesn't need the last nine anyway.
THREE — ACTIONS TO TAKE TODAY
For the investor, price the meter, not the model. Raw intelligence is deflating 79x a year and open weights keep the floor dropping, so margin on tokens compresses from here. The money accrues to the two things that don't commoditize: the harness that routes the work and the verifier that proves it. Underwrite the toll booth and the ruler, not the benchmark score.
For the business owner, build the ruler before you hire the robot. Don't deploy an agent until you've got a cheap, private check that defines "right" for the task, one the model can't see. Then make it quote you 90/95/99% and buy only the nines the job's value justifies. Most of your work isn't 99% work, and paying for nines you don't need is the new way to light money on fire.
For the parent, raise the judge, not the doer. The agent will produce the essay, the code, the analysis for a dime. The scarce, durable skill is the one the machine can't buy its way out of: knowing what "right" looks like and being able to check it. Teach taste, teach verification, teach the nerve to tell a confident machine it's wrong.
FIVE — STORIES TO KEEP YOU INFORMED
Thursday, October 8
Anthropic drops Haiku's price 90% and hands you a dial. (Full analysis above.) Claude Haiku 5.5 runs at $0.10 in and $0.50 out per million tokens, with five effort levels and a computer-use score that jumped from 15.7% to 72.4% in a year. The cheap model is now good enough to be the hands.
Musk's Grok Bot starts routing to Claude. (Full analysis above.) xAI's bot will now pick outside models by task, Claude Opus 5.5 among them, while Anthropic already sits on all the compute at SpaceX's Colossus 1. When your own harness routes to a rival's model, you've admitted the model isn't the moat. The router is.
The open-weight wall goes up the same week the price drops. Mistral Large 4 (over a trillion parameters, weights later this month), Nvidia-backed (NASDAQ: NVDA) Reflection's Beam, and Germany's Aleph Alpha Kolibri all landed, aimed straight at the Chinese open models that lead the field. This is the thing keeping frontier pricing honest.
OpenAI drops 372 math results nobody can check. The company says an agent generated them from roughly one prompt, then withheld the prompts. Mathematicians can't tell what's genuinely new. A breakthrough you can't verify isn't a breakthrough, it's a press release, and it's the whole argument for owning your own ruler.
That $70B OpenAI number is really closer to $50B. Per a CNBC source, the headline ARR folds in revenue-sharing from big partners like Microsoft (NASDAQ: MSFT) and Amazon (NASDAQ: AMZN). Strip the partner plumbing and the real run-rate is about a third smaller. Frontier revenue growth is partly an accounting posture, not pure demand.
— Harry and Anthony
Sources:
Anthropic, Claude Haiku 5.5 launch; EdTech Innovation Hub, "Anthropic says its new Claude Haiku 5.5 delivers more AI capability at 75% lower cost" (Oct 8, 2026); PYMNTS, "AI Answers Get Cheaper as Model Choices Multiply" (Oct 8, 2026)
Help Net Security, "Anthropic's new budget model gets much better at ignoring hidden commands" (Oct 8, 2026), computer-use and prompt-injection figures
Implicator.ai Morning Briefing (Oct 8, 2026), Grok Bot multi-model routing and SpaceX Colossus 1 compute agreement
PYMNTS (Oct 8, 2026), Mistral Large 4, Reflection Beam, Aleph Alpha Kolibri open-weight releases
Scientific American via Implicator.ai (Oct 8, 2026), OpenAI's 372 math results, prompts withheld
Tae Kim (@firstadopter) and CNBC source (Oct 8, 2026), OpenAI ARR $70B versus about $50B ex-partner revenue
CO/AI, Two-Way Street (Oct 5, 2026); I Don't Believe in the No-Win Scenario (Oct 2, 2026)