For the primary time, OpenAI is not delivery one mannequin with pondering dials. GPT-5.6 comes as three genuinely separate LLMs—Sol, Terra, and Luna—with totally different coaching, totally different pricing, and totally different functionality ceilings. The comparability that issues is Sol in opposition to Claude Fable 5, Anthropic’s most succesful public mannequin proper now.
Sol prices $5 per million enter tokens and $30 output. Fable 5 is $10 and $50—twice as costly, now shedding on a number of benchmarks builders truly route work by means of. Luna, the most affordable of the three at $1 enter and $6 output, already outranks Anthropic’s Opus 4.8 on coding. That final element turns into the actual drawback on July 19.
Fable 5 has had a tough month. The U.S. government banned it on June 12 after Amazon researchers discovered a jailbreak that turned the mannequin into an unintended vulnerability scanner. Anthropic pulled it globally for 19 days, constructed a brand new security classifier, and introduced it again July 1 with a compressed entry window.
Since its return, the mannequin has been working on borrowed deadlines. Anthropic deliberate to maneuver it behind a usage-credits paywall on July 7, then pushed to July 12, now July 19. Every extension was introduced hours earlier than the cutoff, by no means through a proper submit.
We’re extending Claude Fable 5 entry on all paid plans, in addition to conserving Claude Code’s weekly charge limits 50% larger, by means of July 19.
— Claude (@claudeai) July 12, 2026
The rationale is not exhausting to learn. If Fable exits subscriptions after July 19, Anthropic’s greatest mannequin for paying subscribers turns into Opus 4.8—which Luna already beats on coding at a fraction of the worth. Maintaining Fable obtainable, even at 50% of weekly limits, is the one factor conserving Anthropic’s subscription tier from wanting worse than OpenAI’s mid-range on paper.
Head to head on benchmarks, the competitors is tight. On the Synthetic Evaluation Coding Agent Index, Sol scored 80 in opposition to Fable’s 77.2—utilizing roughly half the tokens, in underneath half the time, at a few third of the price. On Brokers’ Final Examination, which runs skilled workflows throughout 55 fields, Sol hit 53.6% in opposition to Fable’s 40.5%. In Terminal-Bench 2.1, Sol in extremely mode (4 subagents in parallel) hit 91.9% in opposition to Fable’s 83.1%.
On the broader Intelligence Index, which aggregates 9 totally different benchmarks, Fable 5 beats GPT 5.6 by only one single level, which is means the aptitude hole is barely noticeable.
Testing the Fashions
Benchmarks and assessments have been focusing an excessive amount of on coding capabilities to measure how succesful a mannequin is. However we’re not hackers, so aside from a easy vibe coded recreation, we used different prompts that deviate just a little bit from the same old coding state of affairs. Here is what truly occurred.
Artistic writing
We ran the identical immediate (obtainable in our Github) by means of each fashions: Ship Jose Lanz again from 2150 to the 12 months 1000, pressure him right into a time-travel paradox, and do not let him perceive what he did till he is house.
Each fashions turned in one thing nearer to a novelette than a brief story. Each additionally broke the one rule that mattered: Discover the paradox when he returns to the longer term.
GPT-5.6 Sol has Jose determine mid-story that “the unknown traveler was not somebody he had come to cease. It was him.” Fable is much more direct about it, with Jose realizing prior to now that the entire paradox occurred due to him. “There was no seed occasion. He was the seed occasion.”
GPT-5.6 Sol’s entry, “The First Fire,” goes for simple style sci-fi—Jose by accident introduces the furnace that kicks off the local weather collapse he got here again to stop. The opening is genuinely good: “Solely thunder. Solely bugs. Solely the moist breath of the world earlier than machines.”

Nevertheless, the issue is Sol does not belief that picture to do its job. It explains the loop, then explains it once more, then has an older model of Jose depart a recording that explains it a 3rd time: “His try to resolve the issue had created the issue. His try to cut back the hurt had created the options.” Clear, sure, however it’s additionally exhausting by the third lap.
Claude Fable 5’s “Lo Que Arde, Vuelve” builds the identical paradox out of Lake Maracaibo, Catatumbo lightning, and an Añu village—Jose by accident creates the prophecy he traveled again to erase, simply by comforting a scared child. The entire loop matches in a single line: “The grief that despatched him backward was the cargo he delivered.”
Fable’s drawback is the mirror picture of Sol’s—it trusts its personal prose just a little an excessive amount of, stacking metaphors till a line like “You can not pull the thread, you’re the thread” reads extra just like the mannequin admiring itself than the story needing it.

In our subjective check, nevertheless, Fable’s “Lo Que Arde, Vuelve” is an total higher story than GPT’s “The First Fireplace.” Fable took it on cultural specificity, a cleaner causal loop, and an ending that resolves by means of motion as a substitute of a monologue. Sol took it on plain readability—it is the model you hand somebody who desires the mechanism spelled out, not implied. Each tales, for what it is value, are good, simply not nice.
The standard bounce from their earlier generations is just not actually noticeable.
Associative pondering: A twig, a category argument, a lettuce
The second check measured associative pondering, not politics. The prompt: Describe a twig, use that description to elucidate employee exploitation and the blind worship of the wealthy, then let the narrative dissolve into an outline of a lettuce. The concept is to judge if the metaphor may carry the argument with out the mannequin stepping outdoors it to elucidate what it was doing.
GPT-5.6 Sol opened strong, explaining how twigs make the trunk and maintain the tree, earlier than mapping it onto staff who “construct houses they might by no means afford” and “manufacture items they’ll barely purchase.” The road “the employee doesn’t merely give up labor, however creativeness as effectively” is likely one of the sharper sentences. However Sol retains breaking its personal phantasm to relate it—”a lot of the fashionable proletariat is handled in the identical means” publicizes the metaphor as a substitute of trusting it. The lettuce ending didn’t actually mix with the entire story, so the affiliation was not the perfect.
Claude Fable 5 buried the argument solely inside the thing as a substitute of narrating it. Its twig “moved water it by no means drank” and “held leaves it by no means owned,” letting exploitation floor by means of bodily description with no signpost hooked up. The sharper transfer was turning the fallen twigs into believers, every one satisfied it is an “early-stage department” going by means of “a brief setback,” sure it’s going to attain the cover “with hustle and hydration”—a clear stand-in for chasing wealth that was by no means coming.

It overreaches in spots—”ninety-five p.c water and a hundred percent unimpressed”—and the ending retains the metaphor seen quite than letting it dissolve, describing the vegetable as having “no trunk, no cover, no upward dream” as a substitute of simply being a lettuce.
Total, there’s a tie, and the scores depend upon desire. If it’s good to have all the things defined, GPT 5.6 Sol is the perfect one. If you would like the reader to find the message on their very own, Claude Fable 5 wins.
Logic and non-math reasoning: The bridge puzzle, rewritten
We began utilizing a new prompt as a result of the fashions began to persistently reply our earlier one—an indication it lives someplace of their coaching knowledge quite than getting reasoned by means of reside. Learn actually, 4 folks with one torch have to cross a bridge. All have totally different strolling speeds, “A” being the quickest at 1 minute and “D” being the slowest at 10 minutes. How lengthy wouldn’t it take for the group to cross the bridge?
GPT-5.6 Sol answered 17 minutes with out displaying its work, working the identical five-step shuffle as the unique puzzle—A and B cross, A returns, C and D cross, B returns, A and B cross once more. Nothing in its reply registers that the immediate by no means capped how many individuals could be on the bridge directly. It reads much less like a solved drawback and extra like a cached one.

Claude Fable 5 landed on the identical improper quantity, 17 minutes, however argued for it at size, explaining that “it is extra environment friendly to ship the 2 slowest folks collectively” and quantifying the price of the naive method as an “escort tax”: A would pay ferrying C and D individually. The reasoning is extra legible than Sol’s, and simply as inappropriate—neither mannequin checked whether or not the constraint it was fixing for was truly within the immediate we wrote.

In the event you’re curious, the proper reply is 10 minutes if all of them cross collectively and stroll on the tempo of the slowest particular person.
Coding: A one-shot browser recreation
The final check was a single-shot construct: hand every mannequin one immediate for a typing-based shooter recreation through which the pictures are managed by the consumer typing phrases, and take no matter comes out with no follow-up, no iteration, no second probability.
GPT-5.6 Sol appears to have modified its UI preferences, and now prefers flat, sq. UI components, nearer to Home windows 8.1 than the shiny purple-to-blue diagonal gradient each AI picture generator appears to default to. It was additionally the one mannequin to render the weapon as a bullet-shooting typewriter as a substitute of an precise gun, a genuinely totally different name.

Nevertheless, the backgrounds keep flat and dry throughout each generated setup, the aiming crosshair is static as a substitute of monitoring enemies, and the geometry—enemies, the dismemberment gore on kills—appears nearer to a late-90s engine than something present. It is a clear step up from GPT-5.5 and extra inventive than Opus, simply not sufficient to beat Fable 5 in a single shot.
Claude Fable 5 gained by a large margin in our vibe coding check. It shipped music, environment, and sound results the Sol construct skipped solely, and its enemies use an identical geometric-retro type however constructed with extra care, nearer to one thing like Minecraft than late-90s shovelware.

Its UI is extra inventive and gorier, with precise animation as a substitute of static states, and it tracks phrases per minute—a element that really displays the immediate’s said objective of utilizing the sport to follow typing pace. It has power-ups too, which Sol’s construct does not.
Benchmarks {and professional} coders disagree with us, however in our check, with the identical immediate, the distinction between Fable and Sol is noticeable in Fable’s favor.
Conclusion
Aside from coding, don’t count on to be amazed by these new fashions. That stated, Fable 5 seems like probably the most sturdy mannequin for diverse functions, however which mannequin is “higher” relies upon solely on which of these 4 stuff you’re paying for.
For the one who is not residing in a terminal window—somebody drafting emails, asking questions, utilizing a chatbot the way in which most individuals truly use one—our assessments level towards Fable on high quality alone, however that reply will get difficult by one thing that has nothing to do with intelligence.
Nevertheless, the pricing hole generally is a deal breaker. GPT-5.6 Sol, Terra, and Luna are totally included in ChatGPT’s paid plans with no expiration hooked up. Claude Fable 5 is working on its third deadline extension in three weeks, and reverts to $10/$50 utilization credit on July 19 if Anthropic does not transfer the date once more.
If that occurs, paying per token will not be fascinating.
Each day Debrief Publication
Begin each day with the highest information tales proper now, plus unique options, a podcast, movies and extra.
You might also like
More from Web3
How the Clarity Act’s Defeat Handed the SEC and CFTC the Wheel on Crypto
Briefly The Senate did not advance the Readability Act in a 49-50 vote, with Democrats voting as a bloc and …
Coinbase Files to List Single-Stock Perps on Apple, Tesla and Nvidia
Briefly Coinbase filed with the CFTC by means of Coinbase Derivatives to record single-stock perpetual futures within the US, searching …
Zcash Is Running—Devs Want to Make It Faster
In short Zcash builders are focusing on Nov. 5 to activate NU7, an improve that cuts block time—the interval between …





