Rank #3 / 9
Claude Fable 5.1
Claude Fable 5.1 is Anthropic's Claude model (proprietary). It is ranked 3 of 9 models on the Visual Intelligence Index.
- Anthropic
- Proprietary
- Claude
- Structured output
- Tool calling
- Streaming
- Batch API
Comparison Summary
Claude Fable 5.1 ranks 3 of 9 on the perception index. Immediately ahead is Gemini 3.8 Flash. Immediately behind is Qwen3.8-Max. List price is $10.00 per million input tokens and $50.00 per million output. 3 of 5 priced models cost less to prompt. The context window is 1M. Every score is from our own evaluation runs; list prices are as published by the provider. Strongest capability in this set is Causal Reasoning (79%). Weakest is Action / Event Understanding (74%).
- Visual Intelligence
- 76.2%
- Availability
- Proprietary
- Input / 1M tokens
- $10.00
- Output / 1M tokens
- $50.00
- Context Window
- 1M
- Parameters
- —
All benchmark scores
| Benchmark | Metric | Score | Coverage | Setting |
|---|---|---|---|---|
| Video-MME v2 | Accuracy, no subtitles | 66.2% | 99.7% (3200) | 1 fps |
| LVBench | Test accuracy | 75.0% | Full | 1 fps |
| Perception Test | Overall accuracy | 77.3% | Full | 1 fps |
| NExT-QA | Accuracy (hard split) | 86.8% | Full | 1 fps |
| Q-Bench Video | Overall accuracy | 71.0% | Full | 1 fps |
| EgoSchema | Accuracy (fullset) | 80.4% | 99.8% (500) | 1 fps |
| UCF101-AD | Accuracy | 60.6% | Full | 1 fps |
Benchmark Scores
Compare reported model scores across each available benchmark or capability index.
9 of 9 models
Display
Capabilities Index Scores
Compare reported model scores across each available benchmark or capability index.
9 of 9 models
Display
Usability
3.0/ 5
What it's like to build against this model, scored out of five from the developer-facing capabilities in the dataset.
- Reachable1.0
There is an endpoint you can call without hosting anything.
- First-party API — supported
- Third-party API — supported
- Portable0.0
You can run it yourself, and are not tied to one vendor.
- Open weights — not supported
- Self-hostable — not supported
- Multimodal input0.0
It takes the footage directly, rather than frames you extracted.
- Video ingestion — not supported
- Audio — not supported
- Programmable1.0
Output you can parse, and tools it can call on its own.
- Structured output — supported
- Tool calling — supported
- Operable1.0
Usable interactively and in bulk, not only one call at a time.
- Streaming — supported
- Batch API — supported
Other Anthropic models
Other models from the Anthropic family.
| Date | Model | Visual Intelligence |
|---|---|---|
| Sep 1, 2026 | Claude Fable 5.1 | 76.2% |