The Sand Hill Hour · AI Desk
Sam Mogman comes for King Dario with GPT‑5.6
What's going on here?
The AI-frontier war is so f*ckn back.
(and as if we truly ever got a breather, lmao).
Anywhooooo, you know the drill. Fifth Adderall down, sugar-free Red Bull cracked open, let's get into the juice:
Today, the Mogman himself came for King Dario's crown with the drop of GPT-5.6.
It launched with three new models:
- Sol (the flagship)
- Terra (the middle child)
- Luna (the cheapest tier)
But Sol is the one giving Dario a run for his money…
Sol beat Anthropic's Fable 5 on the big independent AI-coding leaderboard (the Artificial Analysis Coding Agent Index), 80 to 77.2, in half the time with just half the tokens.
Sol's “Ultra Mode” even beat Anthropic's Mythos at completing real tasks inside a computer terminal (like installing software and repairing broken setups), finishing about 92/100 tasks. But OpenAI ran this test itself, so take it with a grain of salt.
Sam announced the release on X last night and told CNBC this morning that Sol is “54% more token-efficient at agentic coding” compared to any other model on the market.
The Sam and Dario beef is sizzling hot right now.
What does this mean?
Early verdict: Sam cooked.
But don't be too quick to crown him the new king. The race is closer than OpenAI wants you to believe:
- Sol scores higher on the Artificial Analysis Coding Agent Index. This is an independent leaderboard. So that means something.
- But Fable 5 still crushes real repo work: hand it actual bugs from real GitHub projects and it fixes 8/10 bugs. Third-party estimates say Sol only fixes about 6.5/10.
- Mythos is also still the best security model. It beats Sol 78 to 73.5 in hacking-skill tests (finding and patching software weaknesses).
The “winner” depends on which scoreboard you look at.
Why should I care?
1) Sol is cheaper, and not worse.
Sol costs about half of what Fable 5 costs: $5 in and $30 out per million tokens, against Fable's $10 and $50. And since it uses fewer tokens to do the same work, your bill shrinks twice. OpenAI also claims that Luna beats Claude Opus 4.8 at $1 and $6. If you run agentic coding at scale, those savings stack up fast.
2) Capability-maxxing vs safety-maxxing.
Two different philosophies are competing right now. While OpenAI gives the public access to its most powerful models, Anthropic withholds public access to its full-power version of Fable (aka Mythos). Even Fable deliberately routes most biology and chemistry requests to an older model for safety.
Channel 940 News The Sand Hill Hour
Rex Remington returns at the top of the hour.