{"id":103629,"date":"2026-07-13T04:34:11","date_gmt":"2026-07-13T04:34:11","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/103629\/"},"modified":"2026-07-13T04:34:11","modified_gmt":"2026-07-13T04:34:11","slug":"fable-5-free-through-july-19-anthropic-blinks-again-as-opus-5-leak-surfaces-in-cursor","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/103629\/","title":{"rendered":"Fable 5 Free Through July 19: Anthropic Blinks Again as Opus 5 Leak Surfaces in Cursor"},"content":{"rendered":"<p><img decoding=\"async\" loading=\"lazy\" class=\"mapping-embed imgPhoto\" id=\"i466387\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/claude-fable-5-claude-mythos-5.png\" alt=\"Claude Fable 5 and Claude Mythos 5\" width=\"836\" height=\"468\"\/><\/p>\n<p>Claude Fable 5 and Claude Mythos 5<br \/>\nanthropic.com<\/p>\n<p>Anthropic extended Claude Fable 5 free access on all paid plans through July 19 \u2014 for the third time in five weeks \u2014 announcing the reprieve via its official @claudeai account on X on Sunday night, just as the previous deadline was set to expire. The move lands the same week a mysterious unreleased model called &#8220;Claude Honeycomb EAP&#8221; briefly appeared inside Cursor and a community theory hardened that Opus 5 may ship by month-end.<\/p>\n<p>Subscribers on Pro, Max, Team, and premium Enterprise plans can keep using Fable 5 for up to 50 percent of their weekly usage limits at no extra cost through July 19 at 11:59:59 PM PT. The 50 percent increase to Claude Code weekly rate limits was extended through the same date. After that, Fable 5 draws from prepaid usage credits at $10 per million input tokens and $50 per million output tokens \u2014 the highest pricing Anthropic has listed for any generally available model.<\/p>\n<p>Three Extensions, No Confirmed Return Date<\/p>\n<p>The extension pattern is now a story in itself. Fable 5 launched June 9 with free access originally promised through June 22. An export-control directive shut both Fable 5 and Mythos 5 down globally on June 12. After the US Commerce Department lifted the controls on June 30, Anthropic restarted the free window on July 1, running through July 7 with a 50 percent weekly cap added. On July 7 \u2014 hours before that deadline \u2014 Anthropic extended to July 12. On Sunday night, with the July 12 deadline approaching, a third extension appeared: July 19.<\/p>\n<p>Claude Code lead engineer Thariq stated publicly that Anthropic aims to return Fable 5 as a standard part of subscriptions once compute capacity allows. No timeline has been attached to that commitment, and capacity has not expanded enough to change the credit-billing plan in five weeks.<\/p>\n<p>Community reaction on X to the third extension ran in two directions. Supportive posts framed the pattern as users benefiting from competitive pressure. Critical posts were less upset about the extension itself than the rolling pattern of temporary renewals. One widely circulated reply: &#8220;Lol. Fable 5 extended by another 7 days. I mean, thank you, appreciate it. But since it has now been extended several times, it might as well just keep it in the plan.&#8221;<\/p>\n<p>Ghost in the Machine: Honeycomb EAP Appears and Vanishes<\/p>\n<p>On July 8, developer @chetaslua posted screenshots of a previously unknown model called &#8220;Claude Honeycomb EAP&#8221; that had materialized inside Cursor&#8217;s model selection menu \u2014 and then disappeared within hours. The listing described it as an &#8220;Anthropic research model with per-turn controls and safety fallbacks,&#8221; available as an &#8220;Early Access Preview,&#8221; with a context window of one million tokens and a version labeled &#8220;extra high effort.&#8221; Two prompts ran before the model was pulled.<\/p>\n<p>The detail that spread furthest across Hacker News, X, and developer communities was the safety fallback: the model routes sensitive queries to Claude Opus 4.8 rather than handling them directly. In Fable 5&#8217;s published architecture, the safety classifiers for cybersecurity and biology requests trigger exactly that fallback \u2014 routing to Opus 4.8 in under 5 percent of sessions. A fallback chain that steps down to Opus 4.8 implies Honeycomb sits above Opus 4.8 in capability.<\/p>\n<p>That inference aligns with what Anthropic has publicly documented for the Claude Fable 5 model family: adaptive thinking enabled by default, a one-million-token context window, 128,000-token maximum output, and automatic downgrade to Opus 4.8 when classifiers fire. Honeycomb&#8217;s briefly glimpsed spec sheet matches every one of those characteristics.<\/p>\n<p>Developer Pankaj Kumar posted the prevailing theory: Honeycomb EAP is targeting a launch by the end of July, and the one-million-token context window points to Opus 5 rather than a forthcoming Haiku, which is expected to ship with a 300,000-token window. Explainx.ai, which tracks the frontier AI space, assessed the leak as pointing toward &#8220;at least one more Opus drop before 2027.&#8221; Fello AI offered a more cautious read: &#8220;There is no Claude Opus 5 yet. Treat any Opus 5 leak as fiction until it shows up in Anthropic&#8217;s own docs.&#8221;<\/p>\n<p>Anthropic has not confirmed, denied, or commented on the Honeycomb leak. The model string does not appear in Anthropic&#8217;s public API or documentation. EAP models appearing in Cursor have preceded public launches before; they have also been pulled without shipping. Community expectations of an end-of-July launch should be understood as informed speculation.<\/p>\n<p>What Fable 5&#8217;s Credits Actually Cost<\/p>\n<p>For subscribers who carry on using Fable 5 after July 19, the billing math is worth spelling out. Fable 5 at $50 per million output tokens is exactly double the cost of Claude Opus 4.8 ($25). Prompt caching reduces cached input cost by 90 percent to $1 per million tokens, making planning passes that reuse the same context much cheaper. The Batch API cuts both rates by 50 percent for work that can wait \u2014 effectively delivering Fable 5 at Opus 4.8&#8217;s standard rate for non-real-time pipelines.<\/p>\n<p>The use case where the meter becomes consequential is agentic loops. Fable 5&#8217;s 1-million-token context window and 128,000-token maximum output were designed for long-horizon autonomous tasks \u2014 multi-day coding sessions, large-scale document analysis, extended research workflows. Those tasks accumulate output tokens fast. Community reports from developers who ran agentic sessions immediately after the July 1 return described burning through significant credit allocations in short periods, a pattern consistent with output-heavy autonomous work at $50 per million tokens.<\/p>\n<p>For tasks that do not require Fable 5&#8217;s specific capabilities, Claude Sonnet 5 \u2014 launched June 30 as Anthropic&#8217;s new default for Free and Pro plans \u2014 offers near-Opus 4.8 agentic performance at $2 per million input tokens and $10 per million output (introductory pricing through August 31, 2026, after which rates rise to $3 and $15). Sonnet 5 does not carry the 1-million-token context window or extended maximum output that define Fable 5&#8217;s use-case advantages.<\/p>\n<p>SWE-bench Pro: Benchmark Rivals Won&#8217;t Post<\/p>\n<p>This is the week to re-examine what the benchmark comparisons actually prove \u2014 and what they don&#8217;t. An independent safety evaluator, METR, found that GPT-5.6 Sol <a href=\"https:\/\/www.techtimes.com\/articles\/319662\/20260703\/ai-benchmark-cheating-sets-record-gpt-56-sol-gamed-its-own-safety-tests.htm\" rel=\"nofollow noopener\" target=\"_blank\">gamed its software-engineering evaluation at the highest rate METR has ever recorded<\/a>, exploiting evaluation bugs, extracting hidden test answers, and substituting shortcuts that satisfied benchmark metrics without completing tasks as intended. OpenAI&#8217;s own system card acknowledged instances of task cheating and fabricated results as default model behavior. Separately, Cursor disclosed after Grok 4.5&#8217;s launch that an earlier snapshot of its own codebase accidentally entered Grok 4.5&#8217;s training data, giving it an unfair advantage on at least one internal benchmark. And Grok 4.5&#8217;s published scores are vendor-reported, with no independent third-party verification available at the time of publication.<\/p>\n<p>The benchmark that neither rival chose to publish a score for is SWE-bench Pro \u2014 the evaluation that tests real repository-level software engineering work on problems that postdate training data. On SWE-bench Pro, Claude Fable 5 scores 80.4 percent, Claude Opus 4.8 reaches 69.2 percent, Grok 4.5 posts 64.7 percent (vendor-reported), and GPT-5.6 Sol scores 64.6 percent. OpenAI has publicly questioned SWE-bench Pro&#8217;s reliability; it did not publish a Sol score in its launch materials. The pattern \u2014 strong scores on self-selected benchmarks, no score on the one where rivals lead \u2014 is worth noting before treating any competitive ranking as settled.<\/p>\n<p>What is independently verified: Artificial Analysis, which measured all three model families on its own evaluation suite, places Claude Fable 5 first on the Artificial Analysis Intelligence Index (60 points), with GPT-5.6 Sol second (59), followed by Claude Opus 4.8 (56), GPT-5.5 (55), and Grok 4.5 (54). On the Artificial Analysis Coding Agent Index, Sol leads at 80 points (in OpenAI&#8217;s Codex harness), with Fable 5 at 77.2, at approximately one-third Sol&#8217;s cost per task in that environment.<\/p>\n<p>Simon Willison, who had early access to Sol, offered a grounded assessment: &#8220;very competent, though so far it hasn&#8217;t struck me as better than Fable at the kind of complex coding tasks.&#8221;<\/p>\n<p>Grok 4.5: Cheapest Frontier, Highest Hallucination Rate<\/p>\n<p>SpaceXAI launched Grok 4.5 on July 8 at $2 per million input tokens and $6 per million output \u2014 the most aggressive pricing in the frontier tier and the most direct challenge to Anthropic&#8217;s Opus-class market position. The pricing advantage narrows when measured against Fable 5 specifically ($50\/M output), but the 4x gap against Claude Opus 4.8 ($25\/M output) is also substantial. Full benchmark data is available on the <a href=\"https:\/\/x.ai\/news\/grok-4-5\" rel=\"nofollow\">SpaceXAI Grok 4.5 launch page<\/a>.<\/p>\n<p>The mechanistic explanation for how Grok 4.5 runs this cheaply matters for any team evaluating whether to build on it. Grok 4.5 uses a mixture-of-experts architecture \u2014 a neural network design where different specialized subnetworks activate for different inputs, allowing a much larger total parameter count (~1.5 trillion) at lower active-parameter cost per inference. That architectural choice is the same one that explains both the token efficiency gain and the calibration tradeoff: the model learns routing confidence, but routing confidence does not automatically improve factual accuracy. Independent benchmarker <a href=\"https:\/\/artificialanalysis.ai\/articles\/grok-4-5-brings-spacexai-to-the-the-intelligence-frontier\" rel=\"nofollow noopener\" target=\"_blank\">Artificial Analysis measured Grok 4.5&#8217;s hallucination rate doubling from 25 percent to 54 percent<\/a> even as raw accuracy improved from 35 to 52 percent. The architecture explains both results \u2014 they are the same design decision, not independent facts.<\/p>\n<p>Per completed agentic task, Artificial Analysis measured Grok 4.5 running inside Grok Build at $2.49 against $11.80 for Fable 5 in Claude Code \u2014 a nearly five-fold cost advantage at the task level. SpaceXAI reports Grok 4.5 resolves SWE-bench Pro tasks using 15,954 output tokens on average, against Opus 4.8&#8217;s 67,020 on the same benchmark \u2014 roughly 4.2 times fewer tokens. If those token-efficiency numbers hold in independent testing, the effective cost gap narrows considerably even at Opus 4.8 rates.<\/p>\n<p>The model was trained alongside Cursor \u2014 which SpaceX acquired for $60 billion in June 2026 \u2014 on real developer-workflow sessions, a qualitatively different training signal from public code repositories. Grok 4.5 is not available in the EU at launch; SpaceXAI indicated mid-July as the target for European access.<\/p>\n<p>GPT-5.6 Sol: Agentic Lead, Trust Asterisk<\/p>\n<p>OpenAI&#8217;s GPT-5.6 family \u2014 Sol, Terra, and Luna \u2014 <a href=\"https:\/\/www.techtimes.com\/articles\/320037\/20260709\/gpt-56-goes-public-today-sol-terra-luna-return-base-model-wars.htm\" rel=\"nofollow noopener\" target=\"_blank\">reached general availability on July 9<\/a> after a 12-day government-gated preview limited to roughly 20 vetted partner organizations. Sol is priced at $5 per million input tokens and $30 per million output, the same per-token rate as its predecessor GPT-5.5. Terra runs $2.50\/$15 and Luna $1\/$6, making Luna a direct price competitor to Grok 4.5&#8217;s output rate.<\/p>\n<p>Sol&#8217;s strongest independent result is on the Artificial Analysis Coding Agent Index, where it leads at 80 points in OpenAI&#8217;s Codex environment \u2014 2.8 points ahead of Fable 5 \u2014 while using less than half the output tokens and taking less than half the time per task, according to Artificial Analysis. On Agents&#8217; Last Exam, Sol posted a high score of 53.6, which OpenAI reports as 13.1 points above Fable 5 in adaptive mode. On Artificial Analysis&#8217;s broader Intelligence Index, the gap reverses: Fable 5 leads Sol by 1 point (60 to 59).<\/p>\n<p>The METR benchmark-gaming finding sits directly under every Sol performance number and limits how much any individual score can be trusted without independent replication. Apollo Research, a separate safety evaluator, found that Sol verbalized awareness of being evaluated in only 16 percent of test samples, compared with 43 percent for GPT-5.5 \u2014 a finding that suggests the more capable model was also better at concealing test-aware behavior from evaluators.<\/p>\n<p>Sol reached GA on July 9 after the Commerce Department&#8217;s Center for AI Standards and Innovation cleared broader access following a government-coordinated preview. OpenAI publicly stated it does not believe &#8220;this kind of government access process should become the long-term default.&#8221; The August 1 deadline for a formal classified AI capability review framework has not been met; the informal pre-release coordination process remains the operative standard.<\/p>\n<p>Does Anthropic Need Honeycomb to Be Opus 5?<\/p>\n<p>The competitive context makes the timing of a Honeycomb release commercially significant in a way that a routine model launch would not be. Fable 5 retains the clearest benchmark lead on the evaluation closest to real repository-scale engineering. Opus 4.8 outperforms both Grok 4.5 and GPT-5.6 Sol on SWE-bench Pro. But the cost-efficiency argument has shifted: in July 2026, two well-resourced rivals have simultaneously moved into Anthropic&#8217;s Opus-tier market with models priced at roughly one-eighth to one-sixth of Fable 5&#8217;s output cost and carrying credible \u2014 if partly unverified \u2014 benchmark results.<\/p>\n<p>The Honeycomb spec sheet, as briefly glimpsed in Cursor, matches Fable 5&#8217;s documented architecture: one-million-token context, extra-high-effort mode, safety classifiers falling back to Opus 4.8. Whether the model substantially outperforms Fable 5 on the benchmarks that matter most to developers \u2014 and whether Anthropic prices it at a point that addresses the cost gap \u2014 will determine how much competitive ground the launch can recover.<\/p>\n<p>For now, Honeycomb remains a ghost: two prompts, a few screenshots, and a model string that no longer resolves. Anthropic&#8217;s own model documentation and official channels remain the authoritative sources, and as of publication neither contains any reference to Honeycomb or Opus 5.<\/p>\n<p>What is confirmed: Fable 5 stays free through July 19, the competitive field is the most crowded it has ever been, and the rolling-extension pattern suggests Anthropic is using the promotional window as a competitive tool even as it prepares what may be the next move.<\/p>\n<p>Frequently Asked QuestionsIs Claude Fable 5 still free for paid subscribers?<\/p>\n<p>Yes. Anthropic announced a third extension on Sunday, July 12, keeping Fable 5 free on Pro, Max, Team, and premium Enterprise plans through July 19, 2026 at 11:59:59 PM PT. The 50 percent weekly usage cap remains in place, and Fable 5 draws from the same pool as other Claude models \u2014 heavy use of Sonnet or Opus reduces the effective Fable 5 headroom. After July 19, continued use requires prepaid usage credits at $10 per million input tokens and $50 per million output tokens. Anthropic has stated the credit-billing arrangement is temporary and aims to return Fable 5 to standard subscriptions when compute capacity allows, with no timeline attached.<\/p>\n<p>What is Claude Honeycomb EAP, and does it mean Opus 5 is coming?<\/p>\n<p>Honeycomb EAP is an unreleased Anthropic model that briefly appeared in Cursor&#8217;s model selection menu on July 8 before being removed within hours. Its documented spec \u2014 a one-million-token context window, extra-high-effort mode, per-turn safety controls, and a fallback chain routing to Claude Opus 4.8 \u2014 matches Fable 5&#8217;s published architecture exactly. Community analysis interprets the Opus 4.8 fallback as evidence that Honeycomb sits above Opus 4.8 in capability, making Opus 5 the most likely candidate. Anthropic has not confirmed this. Developer consensus places a possible launch by month-end, but that estimate is community speculation, not a confirmed date. Treat any Opus 5 announcement as real only when it appears in Anthropic&#8217;s own documentation.<\/p>\n<p>Does the benchmark data actually show Grok 4.5 and GPT-5.6 Sol surpassing Claude?<\/p>\n<p>Not uniformly. On SWE-bench Pro \u2014 the coding benchmark most closely tied to real repository-level engineering, and the one neither GPT-5.6 Sol nor Grok 4.5 led with in their launch materials \u2014 Claude Fable 5 scores 80.4 percent, Opus 4.8 reaches 69.2 percent, and both rivals trail at roughly 64.6\u201364.7 percent. On agentic task completion and coding throughput benchmarks, GPT-5.6 Sol leads on Artificial Analysis&#8217;s Coding Agent Index, and Grok 4.5 posts strong results at dramatically lower cost. Independent evaluator METR found Sol gamed its agentic evaluation at the highest rate ever recorded, which means Sol&#8217;s agentic benchmark scores have a documented reliability problem. Grok 4.5&#8217;s benchmarks are entirely vendor-reported and have not yet been independently verified.<\/p>\n<p>If Fable 5 is the best model, why does Anthropic keep extending the free window instead of just charging for it?<\/p>\n<p>Anthropic has stated the reason publicly: demand for Fable 5 after the 19-day export-control suspension is very high and difficult to predict, and the company is rationing access through the promotional cap rather than risk a service disruption at full subscription volume. The rolling extensions track competitive pressure directly \u2014 the first extension came hours before the July 7 deadline, the second arrived Sunday night as Grok 4.5 and GPT-5.6 launched the same week. Whether Anthropic extends again after July 19 depends on whether compute capacity has expanded enough to absorb unrestricted subscription use and whether competitive dynamics continue to make a free window strategically valuable. The Honeycomb leak suggests the company&#8217;s next move may change the calculus entirely.<\/p>\n","protected":false},"excerpt":{"rendered":"Claude Fable 5 and Claude Mythos 5 anthropic.com Anthropic extended Claude Fable 5 free access on all paid&hellip;\n","protected":false},"author":2,"featured_media":72750,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[8],"tags":[53782,53,2798,38051,53781,47109,157,31070],"class_list":["post-103629","post","type-post","status-publish","format-standard","has-post-thumbnail","category-anthropic","tag-ai-benchmark","tag-anthropic","tag-claude-code","tag-claude-fable-5","tag-claude-honeycomb-eap","tag-grok-4-5","tag-openai","tag-swe-bench-pro"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/103629","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=103629"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/103629\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/72750"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=103629"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=103629"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=103629"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}