Anthropic
Claude Fable 5
Input: USD 10 / 1000000 tokens · Standard · Modality: unspecified
Region: global · Endpoint scope: global · Tier: standard · Billing: Unknown · Prompt bracket (inclusive): Unknown–1000000 tokens · Cache TTL: Unknown seconds · Valid through: Unknown
- Claude API direct global Standard only; excludes Batch, Fast, partner billing, negotiated discounts and regional premiums.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
- US-only inference_geo has a 1.1x multiplier; excluded from this global rate.
Cache write: USD 12.50 / 1000000 tokens · Standard · Modality: unspecified
Region: global · Endpoint scope: global · Tier: standard · Billing: Unknown · Prompt bracket (inclusive): Unknown–1000000 tokens · Cache TTL: 300 seconds · Valid through: Unknown
- Claude API direct global Standard only; excludes Batch, Fast, partner billing, negotiated discounts and regional premiums.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
- US-only inference_geo has a 1.1x multiplier; excluded from this global rate.
Cache write: USD 20 / 1000000 tokens · Standard · Modality: unspecified
Region: global · Endpoint scope: global · Tier: standard · Billing: Unknown · Prompt bracket (inclusive): Unknown–1000000 tokens · Cache TTL: 3600 seconds · Valid through: Unknown
- Claude API direct global Standard only; excludes Batch, Fast, partner billing, negotiated discounts and regional premiums.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
- US-only inference_geo has a 1.1x multiplier; excluded from this global rate.
Cache read: USD 1 / 1000000 tokens · Standard · Modality: unspecified
Region: global · Endpoint scope: global · Tier: standard · Billing: Unknown · Prompt bracket (inclusive): Unknown–1000000 tokens · Cache TTL: Unknown seconds · Valid through: Unknown
- Claude API direct global Standard only; excludes Batch, Fast, partner billing, negotiated discounts and regional premiums.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
- US-only inference_geo has a 1.1x multiplier; excluded from this global rate.
- Cache hits and refreshes: Same duration as the preceding write; no single read TTL is asserted.
Output: USD 50 / 1000000 tokens · Standard · Modality: unspecified
Region: global · Endpoint scope: global · Tier: standard · Billing: Unknown · Prompt bracket (inclusive): Unknown–1000000 tokens · Cache TTL: Unknown seconds · Valid through: Unknown
- Claude API direct global Standard only; excludes Batch, Fast, partner billing, negotiated discounts and regional premiums.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
- US-only inference_geo has a 1.1x multiplier; excluded from this global rate.
- Output tokens as listed; separate reasoning/thinking billing treatment is unknown in this source.
Weight license: Unknown / not verified
Developer: Unknown / not verified · API seller: Anthropic
Checkpoint: Unknown / not verified · Lineage: Unknown / not verified
Price observed: 2026-09-29T04:37:41.581Z · Price source checked: 2026-09-29T04:37:41.581Z · Price freshness: unknown
Weight-license reviewed: Unknown. Open weights do not establish permission for every commercial use; consult the linked terms.
Gating: unknown · Code license: Unknown / not verified · API availability: available (dated observation, not a live service check)
Source, conditions and dates
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: identity · anthropic-claude-fable-5: pricing · anthropic-claude-fable-5: availability · anthropic-claude-fable-5: specifications · Retrieved · Effective: Unknown · Checked: Unknown · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: A
- Cached input is the cache-hit/refresh rate; cache writes cost more. US-only inference is 1.1×.
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: weightAccess · Retrieved · Effective: Unknown · Checked: Unknown · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: Unknown
- The June 9, 2026 system card describes Claude Fable 5 as a general-access configuration with safeguards. General access and discussion of model-weight theft do not establish public weight-download availability.
- Review outcome: unknown. No explicit model-specific weight-availability statement was established in the reviewed source; this is not evidence of non-downloadability. No download URL or weight license is asserted.
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: pricing · anthropic-claude-haiku-4-5: pricing · anthropic-claude-opus-5: pricing · anthropic-claude-sonnet-5: pricing · anthropic-claude-opus-5-5: pricing · anthropic-claude-fable-5-1: pricing · Retrieved · Effective: Unknown · Checked: 2026-09-26T05:34:42.597Z · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: A
- Raw SHA-256: cd36fcc6d6d9e17ac2a8643a3f769e42ec6cf33a8becdbfd98bc3ba8832f1aee; collector anthropic-markdown-2.
- Pricing observation only. Display-name mapping is not official API identity evidence. No new access, availability, weight, retirement or model metadata assertion.
- This page provides detailed pricing information for Anthropic's models and features. All prices are in USD. For the most current pricing information, visit claude.com/pricing.
- The following table shows pricing for all Claude models: 1 Cache hits and refreshes on Claude Fable 5.1 and Claude Mythos 5.1 are priced at 0.025x the base input price. 2 Cache hits and refreshes on Claude Opus 5.5 are priced at 0.05x the base input price. All other models use the standard 0.1x multiplier. 3 The $2/$10 per million input/output token pricing for Claude Sonnet 5, announced at launch as introductory pricing through August 31, 2026, is now the standard price. The previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur. MTok: Million tokens. $5 / MTok is $5 for every million tokens. 5m cache writes: Writing a prompt prefix to the 5-minute prompt cache. 1h cache writes: Writing a prompt prefix to the 1-hour prompt cache. Cache hits and refreshes: Reading a prompt prefix from the prompt cache, which also refreshes it. Limited access: Offered separately, by invitation only, as part of Project Glasswing. For access, contact your Anthropic, AWS, or Google Cloud account team. Retired: May still be available on other cloud platforms. See Model deprecations for more. Claude 4.7 and later models and Claude Mythos Preview use a newer tokenizer that contributes to their improved performance on a wide range of tasks. This tokenizer produces approximately 30% more tokens for the same text. The exact increase depends on the content and workload shape. Claude Sonnet 4.6 and earlier models use the previous tokenizer.
- Prompt caching reduces costs and latency by reusing previously processed portions of your prompt across API calls. Instead of reprocessing the same large system prompt, document, or conversation history on every request, the API reads from cache at a fraction of the standard input price. There are two ways to enable prompt caching: Automatic caching: Add a single cache_control field at the top level of your request. The system automatically manages cache breakpoints as conversations grow. This is the recommended starting point for most use cases. Explicit cache breakpoints: Place cache_control directly on individual content blocks for fine-grained control over exactly what gets cached. Prompt caching uses the following pricing multipliers relative to base input token rates: Cache operation Multiplier Duration 5-minute cache write 1.25x base input price Cache valid for 5 minutes 1-hour cache write 2x base input price Cache valid for 1 hour Cache read (hit) 0.1x base input price (0.025x on Claude Fable 5.1 and Claude Mythos 5.1; 0.05x on Claude Opus 5.5) Same duration as the preceding write Cache write tokens are charged when content is first stored. Cache read tokens are charged when a subsequent request retrieves the cached content. A cache hit costs 10% of the standard input price, which means caching pays off after one cache read for the 5-minute duration (1.25x write), or after two cache reads for the 1-hour duration (2x write). On Claude Fable 5.1 and Claude Mythos 5.1, a cache hit costs 2.5% of the standard input price ($0.25 USD per million tokens). On Claude Opus 5.5, a cache hit costs 5% of the standard input price ($0.20 USD per million tokens). These multipliers stack with other pricing modifiers, including the Batch API discount and data residency. For implementation details, supported models, and code examples, see Prompt caching.
- For Claude 4.6 and later models, specifying US-only inference through the inference_geo parameter incurs a 1.1x multiplier on all token pricing categories, including input tokens, output tokens, cache writes, and cache reads. Global routing (the default) uses standard pricing. This applies to the Claude API (first-party) and Claude Platform on AWS. On Claude in Microsoft Foundry, the same 1.1x multiplier applies to deployments that use the US Data Zone Standard deployment type (see Inference geography). Partner-operated platforms (Bedrock and Google Cloud) have independent regional pricing. See Bedrock and Google Cloud for details. Earlier models do not support the inference_geo parameter and always use standard pricing; requests that include the parameter on these models return a 400 error. For more information, see Data residency.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: pricing · anthropic-claude-haiku-4-5: pricing · anthropic-claude-opus-5: pricing · anthropic-claude-sonnet-5: pricing · anthropic-claude-opus-5-5: pricing · anthropic-claude-fable-5-1: pricing · Retrieved · Effective: Unknown · Checked: 2026-09-27T05:40:23.592Z · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: A
- Raw SHA-256: cd36fcc6d6d9e17ac2a8643a3f769e42ec6cf33a8becdbfd98bc3ba8832f1aee; collector anthropic-markdown-2.
- Pricing observation only. Display-name mapping is not official API identity evidence. No new access, availability, weight, retirement or model metadata assertion.
- This page provides detailed pricing information for Anthropic's models and features. All prices are in USD. For the most current pricing information, visit claude.com/pricing.
- The following table shows pricing for all Claude models: 1 Cache hits and refreshes on Claude Fable 5.1 and Claude Mythos 5.1 are priced at 0.025x the base input price. 2 Cache hits and refreshes on Claude Opus 5.5 are priced at 0.05x the base input price. All other models use the standard 0.1x multiplier. 3 The $2/$10 per million input/output token pricing for Claude Sonnet 5, announced at launch as introductory pricing through August 31, 2026, is now the standard price. The previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur. MTok: Million tokens. $5 / MTok is $5 for every million tokens. 5m cache writes: Writing a prompt prefix to the 5-minute prompt cache. 1h cache writes: Writing a prompt prefix to the 1-hour prompt cache. Cache hits and refreshes: Reading a prompt prefix from the prompt cache, which also refreshes it. Limited access: Offered separately, by invitation only, as part of Project Glasswing. For access, contact your Anthropic, AWS, or Google Cloud account team. Retired: May still be available on other cloud platforms. See Model deprecations for more. Claude 4.7 and later models and Claude Mythos Preview use a newer tokenizer that contributes to their improved performance on a wide range of tasks. This tokenizer produces approximately 30% more tokens for the same text. The exact increase depends on the content and workload shape. Claude Sonnet 4.6 and earlier models use the previous tokenizer.
- Prompt caching reduces costs and latency by reusing previously processed portions of your prompt across API calls. Instead of reprocessing the same large system prompt, document, or conversation history on every request, the API reads from cache at a fraction of the standard input price. There are two ways to enable prompt caching: Automatic caching: Add a single cache_control field at the top level of your request. The system automatically manages cache breakpoints as conversations grow. This is the recommended starting point for most use cases. Explicit cache breakpoints: Place cache_control directly on individual content blocks for fine-grained control over exactly what gets cached. Prompt caching uses the following pricing multipliers relative to base input token rates: Cache operation Multiplier Duration 5-minute cache write 1.25x base input price Cache valid for 5 minutes 1-hour cache write 2x base input price Cache valid for 1 hour Cache read (hit) 0.1x base input price (0.025x on Claude Fable 5.1 and Claude Mythos 5.1; 0.05x on Claude Opus 5.5) Same duration as the preceding write Cache write tokens are charged when content is first stored. Cache read tokens are charged when a subsequent request retrieves the cached content. A cache hit costs 10% of the standard input price, which means caching pays off after one cache read for the 5-minute duration (1.25x write), or after two cache reads for the 1-hour duration (2x write). On Claude Fable 5.1 and Claude Mythos 5.1, a cache hit costs 2.5% of the standard input price ($0.25 USD per million tokens). On Claude Opus 5.5, a cache hit costs 5% of the standard input price ($0.20 USD per million tokens). These multipliers stack with other pricing modifiers, including the Batch API discount and data residency. For implementation details, supported models, and code examples, see Prompt caching.
- For Claude 4.6 and later models, specifying US-only inference through the inference_geo parameter incurs a 1.1x multiplier on all token pricing categories, including input tokens, output tokens, cache writes, and cache reads. Global routing (the default) uses standard pricing. This applies to the Claude API (first-party) and Claude Platform on AWS. On Claude in Microsoft Foundry, the same 1.1x multiplier applies to deployments that use the US Data Zone Standard deployment type (see Inference geography). Partner-operated platforms (Bedrock and Google Cloud) have independent regional pricing. See Bedrock and Google Cloud for details. Earlier models do not support the inference_geo parameter and always use standard pricing; requests that include the parameter on these models return a 400 error. For more information, see Data residency.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: pricing · anthropic-claude-haiku-4-5: pricing · anthropic-claude-opus-5: pricing · anthropic-claude-sonnet-5: pricing · anthropic-claude-opus-5-5: pricing · anthropic-claude-fable-5-1: pricing · Retrieved · Effective: Unknown · Checked: 2026-09-28T04:40:19.422Z · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: A
- Raw SHA-256: cd36fcc6d6d9e17ac2a8643a3f769e42ec6cf33a8becdbfd98bc3ba8832f1aee; collector anthropic-markdown-2.
- Pricing observation only. Display-name mapping is not official API identity evidence. No new access, availability, weight, retirement or model metadata assertion.
- This page provides detailed pricing information for Anthropic's models and features. All prices are in USD. For the most current pricing information, visit claude.com/pricing.
- The following table shows pricing for all Claude models: 1 Cache hits and refreshes on Claude Fable 5.1 and Claude Mythos 5.1 are priced at 0.025x the base input price. 2 Cache hits and refreshes on Claude Opus 5.5 are priced at 0.05x the base input price. All other models use the standard 0.1x multiplier. 3 The $2/$10 per million input/output token pricing for Claude Sonnet 5, announced at launch as introductory pricing through August 31, 2026, is now the standard price. The previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur. MTok: Million tokens. $5 / MTok is $5 for every million tokens. 5m cache writes: Writing a prompt prefix to the 5-minute prompt cache. 1h cache writes: Writing a prompt prefix to the 1-hour prompt cache. Cache hits and refreshes: Reading a prompt prefix from the prompt cache, which also refreshes it. Limited access: Offered separately, by invitation only, as part of Project Glasswing. For access, contact your Anthropic, AWS, or Google Cloud account team. Retired: May still be available on other cloud platforms. See Model deprecations for more. Claude 4.7 and later models and Claude Mythos Preview use a newer tokenizer that contributes to their improved performance on a wide range of tasks. This tokenizer produces approximately 30% more tokens for the same text. The exact increase depends on the content and workload shape. Claude Sonnet 4.6 and earlier models use the previous tokenizer.
- Prompt caching reduces costs and latency by reusing previously processed portions of your prompt across API calls. Instead of reprocessing the same large system prompt, document, or conversation history on every request, the API reads from cache at a fraction of the standard input price. There are two ways to enable prompt caching: Automatic caching: Add a single cache_control field at the top level of your request. The system automatically manages cache breakpoints as conversations grow. This is the recommended starting point for most use cases. Explicit cache breakpoints: Place cache_control directly on individual content blocks for fine-grained control over exactly what gets cached. Prompt caching uses the following pricing multipliers relative to base input token rates: Cache operation Multiplier Duration 5-minute cache write 1.25x base input price Cache valid for 5 minutes 1-hour cache write 2x base input price Cache valid for 1 hour Cache read (hit) 0.1x base input price (0.025x on Claude Fable 5.1 and Claude Mythos 5.1; 0.05x on Claude Opus 5.5) Same duration as the preceding write Cache write tokens are charged when content is first stored. Cache read tokens are charged when a subsequent request retrieves the cached content. A cache hit costs 10% of the standard input price, which means caching pays off after one cache read for the 5-minute duration (1.25x write), or after two cache reads for the 1-hour duration (2x write). On Claude Fable 5.1 and Claude Mythos 5.1, a cache hit costs 2.5% of the standard input price ($0.25 USD per million tokens). On Claude Opus 5.5, a cache hit costs 5% of the standard input price ($0.20 USD per million tokens). These multipliers stack with other pricing modifiers, including the Batch API discount and data residency. For implementation details, supported models, and code examples, see Prompt caching.
- For Claude 4.6 and later models, specifying US-only inference through the inference_geo parameter incurs a 1.1x multiplier on all token pricing categories, including input tokens, output tokens, cache writes, and cache reads. Global routing (the default) uses standard pricing. This applies to the Claude API (first-party) and Claude Platform on AWS. On Claude in Microsoft Foundry, the same 1.1x multiplier applies to deployments that use the US Data Zone Standard deployment type (see Inference geography). Partner-operated platforms (Bedrock and Google Cloud) have independent regional pricing. See Bedrock and Google Cloud for details. Earlier models do not support the inference_geo parameter and always use standard pricing; requests that include the parameter on these models return a 400 error. For more information, see Data residency.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.
Primary source ↗ · vendor-docs · Claims: anthropic-claude-fable-5: pricing · anthropic-claude-haiku-4-5: pricing · anthropic-claude-opus-5: pricing · anthropic-claude-sonnet-5: pricing · anthropic-claude-opus-5-5: pricing · anthropic-claude-fable-5-1: pricing · Retrieved · Effective: Unknown · Checked: 2026-09-29T04:37:41.581Z · Reviewed: Unknown · Review after: Unknown · Freshness: unknown · Confidence: A
- Raw SHA-256: 245b5d0472659075015551126902de479990475fa61deb58c5fc20ef5e51a909; collector anthropic-markdown-2.
- Pricing observation only. Display-name mapping is not official API identity evidence. No new access, availability, weight, retirement or model metadata assertion.
- This page provides detailed pricing information for Anthropic's models and features. All prices are in USD. For the most current pricing information, visit claude.com/pricing.
- The following table shows pricing for all Claude models: 1 Cache hits and refreshes on Claude Fable 5.1 and Claude Mythos 5.1 are priced at 0.025x the base input price. 2 Cache hits and refreshes on Claude Opus 5.5 are priced at 0.05x the base input price. All other models use the standard 0.1x multiplier. 3 The $2/$10 per million input/output token pricing for Claude Sonnet 5, announced at launch as introductory pricing through August 31, 2026, is now the standard price. The previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur. MTok: Million tokens. $5 / MTok is $5 for every million tokens. 5m cache writes: Writing a prompt prefix to the 5-minute prompt cache. 1h cache writes: Writing a prompt prefix to the 1-hour prompt cache. Cache hits and refreshes: Reading a prompt prefix from the prompt cache, which also refreshes it. Limited access: Offered separately, by invitation only, as part of Project Glasswing. For access, contact your Anthropic, AWS, or Google Cloud account team. Retired: May still be available on other cloud platforms. See Model deprecations for more. Claude 4.7 and later models and Claude Mythos Preview use a newer tokenizer that contributes to their improved performance on a wide range of tasks. This tokenizer produces approximately 30% more tokens for the same text. The exact increase depends on the content and workload shape. Claude Sonnet 4.6 and earlier models use the previous tokenizer.
- Prompt caching reduces costs and latency by reusing previously processed portions of your prompt across API calls. Instead of reprocessing the same large system prompt, document, or conversation history on every request, the API reads from cache at a fraction of the standard input price. There are two ways to enable prompt caching: Automatic caching: Add a single cache_control field at the top level of your request. The system automatically manages cache breakpoints as conversations grow. This is the recommended starting point for most use cases. Explicit cache breakpoints: Place cache_control directly on individual content blocks for fine-grained control over exactly what gets cached. Prompt caching uses the following pricing multipliers relative to base input token rates: Cache operation Multiplier Duration 5-minute cache write 1.25x base input price Cache valid for 5 minutes 1-hour cache write 2x base input price Cache valid for 1 hour Cache read (hit) 0.1x base input price (0.025x on Claude Fable 5.1 and Claude Mythos 5.1; 0.05x on Claude Opus 5.5) Same duration as the preceding write Cache write tokens are charged when content is first stored. Cache read tokens are charged when a subsequent request retrieves the cached content. A cache hit costs 10% of the standard input price, which means caching pays off after one cache read for the 5-minute duration (1.25x write), or after two cache reads for the 1-hour duration (2x write). On Claude Fable 5.1 and Claude Mythos 5.1, a cache hit costs 2.5% of the standard input price ($0.25 USD per million tokens). On Claude Opus 5.5, a cache hit costs 5% of the standard input price ($0.20 USD per million tokens). These multipliers stack with other pricing modifiers, including the Batch API discount and data residency. For implementation details, supported models, and code examples, see Prompt caching.
- For Claude 4.6 and later models, specifying US-only inference through the inference_geo parameter incurs a 1.1x multiplier on all token pricing categories, including input tokens, output tokens, cache writes, and cache reads. Global routing (the default) uses standard pricing. This applies to the Claude API (first-party) and Claude Platform on AWS. On Claude in Microsoft Foundry, the same 1.1x multiplier applies to deployments that use the US Data Zone Standard deployment type (see Inference geography). Partner-operated platforms (Bedrock and Google Cloud) have independent regional pricing. See Bedrock and Google Cloud for details. Earlier models do not support the inference_geo parameter and always use standard pricing; requests that include the parameter on these models return a 400 error. For more information, see Data residency.
- Claude 4.6 and later models and Claude Mythos Preview include the full 1M token context window at standard pricing. (A 900k-token request is billed at the same per-token rate as a 9k-token request.) Prompt caching and batch processing discounts apply at standard rates across the full context window.