Claude-fable-5-1 is Anthropic's most capable newly released model, suitable for the most demanding reasoning and long-running agent work. Claude Fable 5.1, while keeping input and output prices the same as Claude Fable 5, reduces cache read prices to one quarter of the original and is more capable for long-running agent programming, multi-step research, and document, spreadsheet, and slide processing. (The model is extremely expensive and not recommended for casual use.)
Pricing
Input Modalities
- Text
- Vision
Output Modalities
- Text
Context length
- 1M tokens
Max output
- 128K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
Anthropic claude-fable-5-1
Pricing$11$55
Cache Read$0.275/M tokens
Web Search$0.01/request
Cache Write$13.75/M tokens
Cache Write 5 Minutes$13.75/M tokens
Cache Write 1 Hour$22/M tokens
Context1M
Max output128K
Latency8.1S
Throughput43.5TPS
Uptime
100.00% uptime 2 days ago
100.00% uptime yesterday
100.00% uptime today
AWS claude-fable-5-1
Pricing$11$55
Cache Read$0.275/M tokens
Web Search$0.01/request
Cache Write$13.75/M tokens
Cache Write 5 Minutes$13.75/M tokens
Cache Write 1 Hour$22/M tokens
Context1M
Max output128K
Latency14.6S
Throughput50.9TPS
Uptime
99.73% uptime 2 days ago
99.74% uptime yesterday
0.00% uptime today
Performance for claude-fable-5-1
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python
