Alternatives

Fable 5 is back, but pricey — here's what else actually works

Fable 5 returned July 1, but it's now the most expensive model Anthropic sells — moving to usage credits after July 19 (window extended) — so cheaper routes still matter. There's no free Fable 5 and no perfect drop-in, but these are the real, usable options, ranked by how well they hold up, with the honest caveats.

Quick summary

  • The value pick is GLM-5.2: open-weight, roughly one-sixth the cost, and level with the closed flagships on third-party coding benchmarks.
  • Opus 4.8 is the closest stay-in-Claude option, and was never part of the suspension.
  • New 'mixture' tools (Hermes MoA, OpenRouter Fusion) try to beat any single model by combining the ones you can access — promising, but the standout numbers are self-reported.
  • Pick by your actual job — Fable 5 is back, but many jobs never needed the priciest model.
Value pickGLM-5.2 (open)
Stay in ClaudeOpus 4.8
ExperimentalMixture-of-agents

The honest answer first

Fable 5 has no free public download, and no single model is a clean one-for-one swap. But several models you can use today get you most of the way — and one of them is genuinely close to the closed flagships, for a fraction of the price.

At a glance

The models you can run today — and where the buzzy orchestration tools actually land.

GLM-5.2★ our pick

Open-weight (MIT) — run it yourself, free.

Coding (FrontierSWE)
74.4
~1/6 the cost · about $1.40 / $4.40 per 1M
available to anyone, today
Opus 4.8Claude

Closest to Fable 5 if you live in Claude.

Coding (FrontierSWE)
75.1
Claude flagship pricing
available — never part of the suspension
GPT-5.5OpenAI

Solid all-rounder in the OpenAI stack.

Coding (FrontierSWE)
72.6
OpenAI standard pricing
available now (GPT-5.6 is gated)

Scores are third-party-reported FrontierSWE (long-horizon coding), higher is better. All three are usable today.

Orchestration routes — combine models you can use (claims self-reported)
Hermes MoA — self-host and tunable; claims about +8% over Opus on its own, unreleased benchmark.
OpenRouter Fusion — a good paid second opinion for one-off hard questions; roughly 3x the cost and slower.
Sakana Fugu — real, but poorly received (slow, pricey, black-box). Watch it; don't bet on it yet.
Read the full orchestration verdict →

The value pick: GLM-5.2

GLM-5.2, from Z.ai (Zhipu), is open-weight (MIT), runs a 1M-token context, and on independent coding benchmarks lands right next to Claude Opus 4.8 and ahead of GPT-5.5 — at around one-sixth the cost. If you lost Fable 5 and want something strong and cheap today, start here.

One caveat: using GLM through its hosted API raises data-handling questions for some teams. Because the weights are open, you can also run it yourself to avoid that.

Closest to Claude: Opus 4.8

If your work is built around Claude — Claude Code, the same prompts and review steps — Opus 4.8 is a solid everyday route at half the token price of Fable 5's post-July-7 credits, and it was never part of the suspension. Keep the same repo and brief so any comparison with Fable 5 stays meaningful.

The mixture route: combine the models you CAN use

A newer idea is getting attention: instead of one model, route a task through several and let an aggregator synthesize the best answer — aiming to beat any single model you can access. Two to watch:

Hermes Agent's Mixture-of-Agents (Nous Research, just shipped) exposes 'MoA presets' as virtual models; it claims to beat Opus 4.8 by about 8% and GPT-5.5 by about 11% — but on its own, not-yet-released benchmark, so treat that as self-reported. OpenRouter's Fusion does something similar server-side. Both still run on the models you already have, and both add cost and latency. A third, Sakana's Fugu, exists but has drawn poor reviews for speed and price.

Pick by your job

  • Long coding / agentic workGLM-5.2 or Opus 4.8; try a mixture route if you want to push past a single model.
  • Cost-sensitive or high-volumeGLM-5.2 (open weights, low cost), or a smaller model with good scaffolding.
  • Stay-in-Claude continuityOpus 4.8.
  • Need the top model on a hard jobFable 5 is back — use it with prompt caching and the Batch API to blunt the premium credit rate.

Want it tailored to your case?

These are the options in general. If you tell our AI advisor what you're actually building and your budget, it'll name the single cheapest setup that works for you — specific models, rough monthly cost, and how to wire it up. Free, and a live preview of Coily.

Need the short answer?

Fable 5 is back worldwide as of July 1 — in-plan on Max and Team Premium, on usage credits for Pro. See the live status, or use GLM-5.2 or the new Sonnet 5 for cheaper work.

Read the brief Fabel 5 spelling guide

FAQ

What's the single best Fable 5 alternative?

For most people, GLM-5.2 — it's open, cheap, and close to the closed flagships on third-party coding benchmarks. If you live in the Claude ecosystem, Opus 4.8 is the smoothest switch.

Is there a free Fable 5?

No. Pages offering a 'free Fable 5' or a download aren't the real model. The closest free-to-run option is an open-weight model like GLM-5.2, which you can host yourself.

Do the 'mixture of agents' tools really beat the top models?

They claim to, but the headline numbers are self-reported on the tools' own benchmarks. They can help, but they still depend on the models you can already access, and they cost more and run slower.

From the Field Guide

Stop guessing which model to use

Weekly verified workflows for the credits era: which model in which seat, what it costs per task, and copy-ready prompts — plus an 8-runbook Starter Pack you can download once and keep. Every claim checked against published pricing, sources shown.

Read the free sample issue See plans — $12/mo · $29 one-time

Sources

This page is independent. Official provider pages are the source of record for access, pricing, and policy.

Opus 5 just reset the math — half price, on your plan. One free email on what actually changed.