Anthropic's top tier
Fable sits above Opus in Anthropic's lineup. Use it where a wrong answer costs far more than the tokens.
The Claude Fable 5.1 API: Anthropic's top-tier model, 65% below list price.
claude-fable-5-1from openai import OpenAIclient = OpenAI( base_url="https://api.zurelay.com/v1", api_key="YOUR_ZURELAY_KEY",)stream = client.chat.completions.create( model="claude-fable-5-1", messages=[{"role": "user", "content": "Explain quicksort in two sentences."}], stream=True,)for chunk in stream: if chunk.choices: print(chunk.choices[0].delta.content or "", end="")Base URL https://api.zurelay.com/v1. Works with any OpenAI SDK, and the Anthropic SDK and Claude Code.
Pricing
Pay as you go from prepaid credit, with no subscription and no minimum. Requests that fail are never billed.
| Rate | zurelay | Anthropic list | You save |
|---|---|---|---|
Input per 1M tokens | $3.50 | 65%off | |
Output per 1M tokens | $17.50 | 65%off | |
Cached input per 1M tokens | $0.35 | — | — |
One rate at every prompt length, up to the full 1M context window. Streaming, tool calls and structured outputs cost nothing extra.
Move the sliders to your monthly usage.
$7,800 a year
Overview
Claude Fable 5.1 is Anthropic's current Fable model, from the tier above Opus, built for demanding reasoning and long-horizon agentic work. The zurelay Claude Fable 5.1 API serves it through OpenAI-compatible and Anthropic-compatible endpoints at 65% below Anthropic's list price, so your hardest jobs run for less.
Claude Fable 5.1 is Anthropic's current Fable model, released on September 1, 2026 as the successor to Claude Fable 5. Fable belongs to Anthropic's Mythos class, which Anthropic places above Opus in capability. Anthropic recommends it for demanding reasoning and long-horizon agentic work, or for tasks where Claude Opus at higher effort still falls short on your evals. On zurelay, the Claude Fable 5.1 API uses the model ID claude-fable-5-1 and costs $3.50 per 1M input tokens and $17.50 per 1M output tokens.
Anthropic reports the largest gains over Fable 5 in six areas: agentic coding over long sessions, work with documents, spreadsheets and slides, multistep web research, vision on dense charts and tables in PDFs, reasoning across the full 1M-token window, and computer use. The gap is widest at higher effort. In Anthropic's launch benchmarks, Fable 5.1 scores above both Fable 5 and Claude Opus 5 on the agentic coding, computer use, research and knowledge-work evaluations Anthropic published.
The specs: a 1M-token context window, up to 128K output tokens per request, text and image input, and text output, with a reliable knowledge cutoff of June 2026. Adaptive thinking is always on. Effort runs from low through medium, high (the default) and xhigh to max, and Anthropic says Fable 5.1 at low or medium effort matches or beats Fable 5 at a much lower cost. List prices per token match Fable 5, but cache reads list at a quarter of Fable 5's rate, which Anthropic estimates makes typical workloads about 25% cheaper.
Moving from Fable 5 takes a few checks. Forced tool choice now returns an error, so keep tool_choice on auto and use strict tool schemas or structured outputs when you need guaranteed JSON. If you pass thinking blocks back through the Messages API, keep the history append-only, because editing earlier turns invalidates them. In agent loops, Fable 5.1 may make one tool call per turn where Fable 5 batched several; Anthropic suggests a one-line instruction to batch independent calls.
Strengths
Fable sits above Opus in Anthropic's lineup. Use it where a wrong answer costs far more than the tokens.
Anthropic reports its biggest gains in long sessions: multi-file features, large refactors and migrations, debugging and code review.
Higher accuracy on multistep web research, and stronger work that runs from a first question to a finished report, spreadsheet or slide deck.
Cache reads list at a quarter of Fable 5's rate, so agents that re-read a large cached prefix cost much less per turn.
Use cases
Hand off multi-file features, migrations and code review that run for hours, and stream progress back to your CLI, IDE or CI job.
Run multistep research through your own search and fetch tools, where the model follows up on what it finds before it answers.
Read screenshots of dense charts, financial filings and nested tables, and return the numbers as structured JSON.
Load a large codebase or document set into the 1M-token window and ask questions that connect details across all of it.
Get started
No waitlist and no new SDK. If your code already talks to OpenAI, it already talks to zurelay.
Sign up, add credit and create an API key. Set a monthly budget or a rate limit per key if you like.
Change the base URL. Everything else in your code stays the same.
https://api.zurelay.com/v1Send chat completions as usual. Streaming, tool calls and usage reporting work as you expect.
claude-fable-5-1Works with the tools you already use
Compare
$1.40 / $7.00 per 1M tokens
Pick Claude Opus 5.5 for most workloads at a lower price; Anthropic suggests starting there and moving up to Fable 5.1 only when your evals fall short.
Claude Opus 5.5 API$3.50 / $17.50 per 1M tokens
Pick Claude Fable 5 if your code relies on forced tool choice or you need results that match evals already run on it.
Claude Fable 5 API$0.80 / $4.00 per 1M tokens
Pick Claude Sonnet 5.5 for fast, well-scoped tasks where latency and cost matter more than peak reasoning.
Claude Sonnet 5.5 APIFAQ
Anything else, write to sales@zurelay.com and an engineer will answer.
zurelay charges $3.50 per 1M input tokens and $17.50 per 1M output tokens, with cached input at $0.35. Anthropic's list price is $10.00 input and $50.00 output, so you save 65%. The rate is the same at every prompt length, up to the full 1M-token window.
Yes. zurelay serves claude-fable-5-1 at 65% below Anthropic's list price, paid from prepaid credit with no subscription. It is the same model, not a smaller substitute. You keep your code and change the base URL and key.
Create a zurelay API key, set the base URL to https://api.zurelay.com/v1 and set the model to claude-fable-5-1. Chat completions, streaming and tool calls use the request shapes you know from OpenAI. Keep tool_choice on auto, since forcing a tool (required, or a named function) returns an error on this model, and leave temperature and top_p unset.
Yes. Set the Anthropic SDK's base URL, or ANTHROPIC_BASE_URL for Claude Code, to https://api.zurelay.com with no /v1, because the client adds /v1/messages itself. Authenticate with your zurelay key and choose claude-fable-5-1 as the model. Anthropic notes that Claude Code runs Fable 5.1 at high effort by default.
Claude Fable 5.1 has a 1M-token context window, roughly 555k English words on Anthropic's current tokenizer, and returns up to 128K output tokens per request. Thinking counts toward the output limit, so set a large max_tokens at high effort and stream long responses.
Anthropic built it for demanding reasoning and long-horizon agentic work: coding sessions that run for hours, multistep research, document, spreadsheet and slide work, vision on dense charts, and computer use. For routine traffic, Claude Opus 5.5 or Claude Sonnet 5.5 usually costs less per task.
The context window, output limit and list price per token are the same, and cache reads list at a quarter of the old rate. Three API changes can break code: forced tool choice returns an error, earlier models can't read Fable 5.1 thinking blocks, and editing earlier turns invalidates thinking blocks. Anthropic also notes fewer parallel tool calls in some agent loops and less markdown formatting in chat.
Yes. Streaming works on chat completions and on the Messages API; on chat completions, set stream_options.include_usage to get token counts at the end of the stream. Hard tasks at high effort can run long, so stream those requests. Each key can have its own monthly budget and requests-per-minute cap, and failed requests are retried on another route and never billed.
Yes. Your requests run on Claude Fable 5.1 itself, and zurelay never swaps in a different or smaller model. Output is sampled, so exact wording varies between calls on any API. zurelay is an independent service and is not affiliated with or endorsed by Anthropic.
Create a key in seconds, point the SDK you already use at zurelay, and every request costs up to 85% less from the first token.