Anthropic
Operational· 100% uptime, 7 days

Claude Fable 5.1 API

The Claude Fable 5.1 API: Anthropic's top-tier model, 65% below list price.

Price per 1M tokens65%off
Input
$3.50$10.00
Output
$17.50$50.00
Cached input
$0.35
Model IDclaude-fable-5-1
chat.py
from openai import OpenAI
client = OpenAI(
base_url="https://api.zurelay.com/v1",
api_key="YOUR_ZURELAY_KEY",
)
stream = client.chat.completions.create(
model="claude-fable-5-1",
messages=[{"role": "user", "content": "Explain quicksort in two sentences."}],
stream=True,
)
for chunk in stream:
if chunk.choices:
print(chunk.choices[0].delta.content or "", end="")

Base URL https://api.zurelay.com/v1. Works with any OpenAI SDK, and the Anthropic SDK and Claude Code.

Context window
1M tokens
Max output
128K tokens
Input
Text, Image
Output
Text
Released
Sep 1, 2026
Uptime, 7 days
100%

Pricing

Claude Fable 5.1 API pricing. A fraction of the list price.

Pay as you go from prepaid credit, with no subscription and no minimum. Requests that fail are never billed.

RatezurelayAnthropic listYou save
Input
per 1M tokens
$3.50$10.0065%off
Output
per 1M tokens
$17.50$50.0065%off
Cached input
per 1M tokens
$0.35——

One rate at every prompt length, up to the full 1M context window. Streaming, tool calls and structured outputs cost nothing extra.

Savings calculator

Move the sliders to your monthly usage.

Input tokens per month50M
Output tokens per month10M
zurelay
$350.00
Anthropic list price
$1,000
You save every month$650.00

$7,800 a year

Overview

What is Claude Fable 5.1?

Claude Fable 5.1 is Anthropic's current Fable model, from the tier above Opus, built for demanding reasoning and long-horizon agentic work. The zurelay Claude Fable 5.1 API serves it through OpenAI-compatible and Anthropic-compatible endpoints at 65% below Anthropic's list price, so your hardest jobs run for less.

Claude Fable 5.1 is Anthropic's current Fable model, released on September 1, 2026 as the successor to Claude Fable 5. Fable belongs to Anthropic's Mythos class, which Anthropic places above Opus in capability. Anthropic recommends it for demanding reasoning and long-horizon agentic work, or for tasks where Claude Opus at higher effort still falls short on your evals. On zurelay, the Claude Fable 5.1 API uses the model ID claude-fable-5-1 and costs $3.50 per 1M input tokens and $17.50 per 1M output tokens.

Anthropic reports the largest gains over Fable 5 in six areas: agentic coding over long sessions, work with documents, spreadsheets and slides, multistep web research, vision on dense charts and tables in PDFs, reasoning across the full 1M-token window, and computer use. The gap is widest at higher effort. In Anthropic's launch benchmarks, Fable 5.1 scores above both Fable 5 and Claude Opus 5 on the agentic coding, computer use, research and knowledge-work evaluations Anthropic published.

The specs: a 1M-token context window, up to 128K output tokens per request, text and image input, and text output, with a reliable knowledge cutoff of June 2026. Adaptive thinking is always on. Effort runs from low through medium, high (the default) and xhigh to max, and Anthropic says Fable 5.1 at low or medium effort matches or beats Fable 5 at a much lower cost. List prices per token match Fable 5, but cache reads list at a quarter of Fable 5's rate, which Anthropic estimates makes typical workloads about 25% cheaper.

Moving from Fable 5 takes a few checks. Forced tool choice now returns an error, so keep tool_choice on auto and use strict tool schemas or structured outputs when you need guaranteed JSON. If you pass thinking blocks back through the Messages API, keep the history append-only, because editing earlier turns invalidates them. In agent loops, Fable 5.1 may make one tool call per turn where Fable 5 batched several; Anthropic suggests a one-line instruction to batch independent calls.

Strengths

Where Claude Fable 5.1 shines. And what teams build with it.

01

Anthropic's top tier

Fable sits above Opus in Anthropic's lineup. Use it where a wrong answer costs far more than the tokens.

02

Hours-long agentic coding

Anthropic reports its biggest gains in long sessions: multi-file features, large refactors and migrations, debugging and code review.

03

Research and knowledge work

Higher accuracy on multistep web research, and stronger work that runs from a first question to a finished report, spreadsheet or slide deck.

04

Cheaper cached context

Cache reads list at a quarter of Fable 5's rate, so agents that re-read a large cached prefix cost much less per turn.

Use cases

  • Autonomous coding agents

    Hand off multi-file features, migrations and code review that run for hours, and stream progress back to your CLI, IDE or CI job.

  • Deep research

    Run multistep research through your own search and fetch tools, where the model follows up on what it finds before it answers.

  • Dense charts and filings

    Read screenshots of dense charts, financial filings and nested tables, and return the numbers as structured JSON.

  • Whole-repo and corpus questions

    Load a large codebase or document set into the 1M-token window and ask questions that connect details across all of it.

Get started

Call Claude Fable 5.1 in three steps.

No waitlist and no new SDK. If your code already talks to OpenAI, it already talks to zurelay.

  1. 1

    Create a key

    Sign up, add credit and create an API key. Set a monthly budget or a rate limit per key if you like.

  2. 2

    Point your SDK at zurelay

    Change the base URL. Everything else in your code stays the same.

    https://api.zurelay.com/v1
  3. 3

    Use claude-fable-5-1

    Send chat completions as usual. Streaming, tool calls and usage reporting work as you expect.

    claude-fable-5-1

FAQ

Claude Fable 5.1 API questions.

Anything else, write to sales@zurelay.com and an engineer will answer.

How much does the Claude Fable 5.1 API cost on zurelay?

zurelay charges $3.50 per 1M input tokens and $17.50 per 1M output tokens, with cached input at $0.35. Anthropic's list price is $10.00 input and $50.00 output, so you save 65%. The rate is the same at every prompt length, up to the full 1M-token window.

Is there a cheaper Claude Fable 5.1 API than Anthropic's?

Yes. zurelay serves claude-fable-5-1 at 65% below Anthropic's list price, paid from prepaid credit with no subscription. It is the same model, not a smaller substitute. You keep your code and change the base URL and key.

How do I call Claude Fable 5.1 with the OpenAI SDK?

Create a zurelay API key, set the base URL to https://api.zurelay.com/v1 and set the model to claude-fable-5-1. Chat completions, streaming and tool calls use the request shapes you know from OpenAI. Keep tool_choice on auto, since forcing a tool (required, or a named function) returns an error on this model, and leave temperature and top_p unset.

Can I use Claude Fable 5.1 with the Anthropic SDK or Claude Code?

Yes. Set the Anthropic SDK's base URL, or ANTHROPIC_BASE_URL for Claude Code, to https://api.zurelay.com with no /v1, because the client adds /v1/messages itself. Authenticate with your zurelay key and choose claude-fable-5-1 as the model. Anthropic notes that Claude Code runs Fable 5.1 at high effort by default.

What is the Claude Fable 5.1 context window?

Claude Fable 5.1 has a 1M-token context window, roughly 555k English words on Anthropic's current tokenizer, and returns up to 128K output tokens per request. Thinking counts toward the output limit, so set a large max_tokens at high effort and stream long responses.

What is Claude Fable 5.1 best at?

Anthropic built it for demanding reasoning and long-horizon agentic work: coding sessions that run for hours, multistep research, document, spreadsheet and slide work, vision on dense charts, and computer use. For routine traffic, Claude Opus 5.5 or Claude Sonnet 5.5 usually costs less per task.

What changed from Claude Fable 5 to Claude Fable 5.1?

The context window, output limit and list price per token are the same, and cache reads list at a quarter of the old rate. Three API changes can break code: forced tool choice returns an error, earlier models can't read Fable 5.1 thinking blocks, and editing earlier turns invalidates thinking blocks. Anthropic also notes fewer parallel tool calls in some agent loops and less markdown formatting in chat.

Does the zurelay Claude Fable 5.1 API support streaming, and what are the limits?

Yes. Streaming works on chat completions and on the Messages API; on chat completions, set stream_options.include_usage to get token counts at the end of the stream. Hard tasks at high effort can run long, so stream those requests. Each key can have its own monthly budget and requests-per-minute cap, and failed requests are retried on another route and never billed.

Is Claude Fable 5.1 on zurelay the same model as Anthropic's?

Yes. Your requests run on Claude Fable 5.1 itself, and zurelay never swaps in a different or smaller model. Output is sampled, so exact wording varies between calls on any API. zurelay is an independent service and is not affiliated with or endorsed by Anthropic.

More models on zurelay

All models

Stop paying list price.
Start saving today.

Create a key in seconds, point the SDK you already use at zurelay, and every request costs up to 85% less from the first token.