Skip to contentKumoDocs
Sections
On this page
Models Anthropic

Claude Fable 5

Anthropic Claude Fable 5 text model. Available over the Anthropic Messages and Chat Completions protocols.

View as Markdown

Overview

Context

1Mtokens

Max output

128Ktokens

Input

6 USDper 1M tokens

Output

30 USDper 1M tokens

Released

7 June 2026

Model ID
anthropic/claude-fable-5

Start with Claude Fable 5

The model name is already filled in. Mint a key in the console, put it in an environment variable, and the call below runs as it stands — provided the organization’s wallet holds funds: a call with nothing to pay with answers 402.

curl https://api.kumorouter.com/v1/chat/completions \
  -H "Authorization: Bearer $KUMO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "anthropic/claude-fable-5",
  "max_tokens": 1024,
  "messages": [
    {
      "role": "user",
      "content": "Explain tokens in one line."
    }
  ]
}'

The Anthropic-shaped call

curl https://api.kumorouter.com/v1/messages \
  -H "x-api-key: $KUMO_API_KEY" \
  -H "anthropic-version: 2023-06-01" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "anthropic/claude-fable-5",
  "max_tokens": 1024,
  "messages": [
    {
      "role": "user",
      "content": "Explain tokens in one line."
    }
  ]
}'

OpenAI — API parameters

Authorization: Bearer $KUMO_API_KEY

ParameterTypeDefault / rangeDescription
modelstringRequiredPublished model identifier.
messagesarray1–512Conversation messages in chronological order.
max_tokensinteger≥ 1Optional. A request stating neither max_tokens nor max_completion_tokens is answered under the surface default of 32768 output tokens. Model limits also apply.
temperaturenumber0–2Sampling randomness. Omission leaves the choice to the model.
top_pnumber0–1Nucleus sampling probability mass. Omission leaves the choice to the model.
stoparray≤ 4Array of strings that stop generation.
max_completion_tokensinteger≥ 1Optional alternative to max_tokens. Stating both with DIFFERENT values is refused: a precedence rule would silently discard one of two numbers the caller deliberately wrote.
streambooleanfalseStream response events using Server-Sent Events.
stream_optionsobject{include_usage: boolean}Only with stream: true. include_usage adds a final usage event before [DONE]; its default is false.
toolsarray≤ 128Function declarations. parameters carries the original JSON Schema object, up to 65536 bytes per tool, including nested objects and arrays.
tool_choicestring | objectauto | none | required | {type: "function", function: {name}}auto leaves the choice to the model; none forbids a call while the declarations stay visible; required requires a call to some declared tool; the object form requires the named one. Omission leaves the choice to the model.

One completion per request. The gateway rejects frequency_penalty, presence_penalty, seed, logprobs, top_logprobs and logit_bias: each of them would change the generation or the shape of the reply, and accepting one silently would misdescribe the answer you got. n is accepted only at 1 — the protocol default and exactly what the gateway does — and an n above one is rejected. user is accepted but is not forwarded to the model, not stored and not echoed back. The same applies to strict on a function declaration and name on a tool-role message: both are accepted and not forwarded. reasoning_effort is accepted and not forwarded either: no supplier request on this platform carries a reasoning budget, so the answer comes back at the default of the model itself whatever the field says. cache_control on a message, on a content part or on a replayed tool call is accepted and not forwarded either: the supplier wire this protocol is served by has no member for a prompt-cache anchor, so the cached-token counts come back as though the anchor had not been sent. messages[].content is accepted both as a string and as a list of parts: text parts are joined in order into one text, so a request written in parts and the same request written as a string are one request. A part of kind image_url, input_audio, file or refusal is rejected naming its own type: no multimodal capability is verified on this platform for any pair, and a replayed refusal would reach the supplier as ordinary assistant speech. The developer role is accepted and forwarded as system — it is the newer name this protocol gives the instruction role. stop is array-only; tool_choice is the string mode auto, none or required, or an object naming a declared tool. json_object is refused by name: promising valid JSON says nothing about its shape.

Anthropic — API parameters

x-api-key: $KUMO_API_KEY

ParameterTypeDefault / rangeDescription
modelstringRequiredPublished model identifier.
messagesarray1–512Conversation messages in chronological order.
max_tokensinteger1–1048576Required output token ceiling, subject to model limits.
temperaturenumber0–1Sampling randomness. Omission leaves the choice to the model.
top_pnumber0–1Nucleus sampling probability mass. Omission leaves the choice to the model.
stop_sequencesarray≤ 16Array of strings that stop generation.
systemstring | array—System instruction as a string or text blocks.
top_kinteger1–1048576Keep the K most likely tokens when sampling. Forwarded unchanged; omission is not replaced by zero.
streambooleanfalseStream response events using Server-Sent Events.
toolsarray≤ 128Tool declarations. input_schema carries the original JSON Schema object, up to 65536 bytes per tool.
tool_choiceobject{type: auto | any | none | tool}auto leaves the choice to the model; any requires a call; none forbids calls; tool selects a named tool. name is required only for tool and forbidden in other modes.

Uses the native Messages format. OpenAI fields frequency_penalty, presence_penalty, n, seed and response_format do not belong to it. metadata, thinking, output_config and context_management are accepted but not forwarded to the model and are outside verified support. The same applies to tool declaration fields eager_input_streaming, strict, defer_loading, allowed_callers, input_examples and max_uses; server-side tools are unsupported. Tool input_schema is carried as the original JSON Schema object. cache_control: {type: "ephemeral"} is supported on text blocks and tool declarations; cache availability depends on the model.

Identifiers

Canonical name
anthropic/claude-fable-5
Aliases
claude-fable-5
Vendor code
anthropic
Protocols
anthropic_messages, chat_completions
Modalities
text_generation

Capabilities

Protocol and modalityStreamingToolsStructured outputStrict semanticsStructured streaming
Anthropic Messages text generationYesYesUnverifiedUnverifiedUnverified
Chat Completions text generationYesUnverifiedUnverifiedUnverifiedUnverified
Chat Completions text generationYesYesUnverifiedUnverifiedUnverified
Chat Completions text generationUnverifiedYesUnverifiedUnverifiedUnverified

A row describes one protocol-and-modality pair rather than the model as a whole: a capability is proved on a particular surface, and the answer beside it may differ.

How it compares

ModelContextMax outputInput $/MOutput $/MModalities
Claude Fable 51M128K6 USD30 USDtext generation
Claude Sonnet 51M128K1.2 USD6 USDtext generation
Claude Opus 4.81M128K3 USD15 USDtext generation
Claude Fable 5.11M128K6 USD30 USDtext generation
Claude Haiku 4.5200K64K0.6 USD3 USDtext generation
Claude Opus 4.5200K64K3 USD15 USDtext generation
Claude Opus 4.61M128K3 USD15 USDtext generation
Claude Opus 4.71M128K3 USD15 USDtext generation
Claude Opus 5.51M128K2.2 USD11 USDtext generation
Claude Opus 51M128K3 USD15 USDtext generation
Claude Sonnet 4.51M64K1.8 USD9 USDtext generation
Claude Sonnet 4.61M128K1.8 USD9 USDtext generation
Gemini 3 Flash Preview1M65.5K0.3 USD1.8 USDtext generation
Gemini 3.1 Flash Lite1M65.5K0.2 USD0.9 USDtext generation
Gemini 3.6 Flash High1M65.5K0.5 USD2.3 USDtext generation
Gemini 3.7 Flash High1M65.5K0.5 USD2.3 USDtext generation
Gemini 3.8 Flash High1M65.5K0.5 USD2.3 USDtext generation
Codex Auto Review1.1M128K3 USD18 USDtext generation
GPT-5.3 Codex Spark128K32K1.1 USD8.4 USDtext generation
GPT-5.51.1M128K3 USD18 USDtext generation
GPT-5.6 Luna1.1M128K0.1 USD0.7 USDtext generation
GPT-5.6 Sol1.1M128K2.4 USD12 USDtext generation
GPT-6 Astra1.1M128K6 USD30 USDtext generation
GPT-5.6 Terra1.1M128K1.2 USD7.2 USDtext generation
GPT-6 Sol1.1M128K1.1 USD5.5 USDtext generation
Grok 4.5500K500K1.2 USD3.6 USDtext generation
Grok 4.6500K500K1.2 USD3.6 USDtext generation
Grok 4.7500K500K1.1 USD3.3 USDtext generation
GLM-5.21M131.1K0.8 USD2.6 USDtext generation
GLM-5.31M131.1K0.8 USD2.6 USDtext generation

Prices

DimensionRate
Input tokens6 USD per 1M tokens
Cache write7.5 USD per 1M tokens
Cache read0.6 USD per 1M tokens
Output tokens30 USD per 1M tokens

The whole price list →

Integrations

Every link opens its own guide with this model already chosen.

Errors Rate limits Models

Claude Fable 5 Kumo