Compare models
Choose two or three models, including exact thinking and fast versions. Copy the URL to share your selection.
| Comparison | Claude Fable 5.1anthropic/claude-fable-5-1 | Choose model 2 |
|---|---|---|
| Type | Chat | — |
| TokenBaazar price · 50% off | ₹480.00 in / ₹2400.00 out per 1M tokens | — |
| Market price | ₹960.00 in / ₹4800.00 out per 1M tokens | — |
| Context window (provider) | 1,000,000 tokens | — |
| Input token limit (provider) | Within the context window | — |
| Maximum output | 128,000 tokens | — |
| Knowledge cutoff | Jun 2026 | — |
| Vendor input modalities | Text, Image | — |
| Vendor output modalities | Text | — |
| TokenBaazar endpoints | POST /v1/chat/completions · POST /v1/messages | — |
| TokenBaazar input | Text, Image/PDF content through the chat bridge | — |
| TokenBaazar output | Text | — |
| Benchmark evidence | 33 source-linked results | — |
| SWE-bench Pro | 81.2% Provider model evaluated with adaptive thinking at max effort; average over five trials. This is not the catalog default medium-effort configuration. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| DeepSWE v1.1 | 67.4% Provider model evaluated with adaptive thinking at max effort; average over five trials. This is not the catalog default medium-effort configuration. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| SWE-bench Multilingual | 89.1% Provider model evaluated with adaptive thinking at max effort; average over five trials. This is not the catalog default medium-effort configuration. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| SWE-bench Multimodal | 54.7% Provider model evaluated with adaptive thinking at max effort; average over five trials. This is not the catalog default medium-effort configuration. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| CursorBench 3.2.0 | 73.4% System card section 8.8 reports Cursor’s production agent harness at max effort; historical 3.2.0 test, not CursorBench 4.0. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| ProgramBench | 87.6% Section 8.11.1: 166 filtered tasks; hidden test pass rate, mini-swe-agent without the six-hour time limit; max effort. Not interchangeable with the unfiltered public evaluation. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| FrontierSWE v2 | 56.53% Published current mean reward 0.5653 over 170 runs, displayed as a percentage; not the older supplied 56.3 result. Effort is not established by this leaderboard. Proximal · FrontierSWE public leaderboard | — |
| LiveCodeBench (Vals) | 90.52% Vals evaluation: max compute effort, temperature 1, maximum output 128,000; reported uncertainty ±0.86 percentage points. Not the default medium version. Vals AI · Claude Fable 5.1 evaluation | — |
| CursorBench 4.0 | 51.8% Cursor’s live table explicitly lists Fable 5.1 Max at 51.8; Extra High is a separate 51.6 result. Max must not be equated with the catalog xhigh variant. CursorBench public evaluation51.6% Row explicitly labeled 'Fable 5.1 Extra High' (xhigh) at 51.6%, $13.01/task, confirmed directly on Cursor's live table. Distinct from the already-recorded 51.8% 'Fable 5.1 Max' row and the 49.2% 'Fable 5.1 High' row on the same leaderboard — effort levels are not interchangeable. Cursor · CursorBench 4.0 public leaderboard | — |
| ARC-AGI-1 | 97.5% Section 8.16: ARC Prize verified semi-private dataset, max effort. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| ARC-AGI-2 | 90% Section 8.16: ARC Prize verified semi-private dataset, max effort. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| GMMLU | 94% Section 8.18.1: 42 languages, adaptive thinking max effort, single trial, no tools or custom system prompt. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| MILU | 93% Section 8.18.2: 11 languages, adaptive thinking max effort, single trial. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| HealthBench (raw) | 66.7% Section 8.17: max effort, five trials, Opus 4.8 grader; safety classifiers and Opus 5 refusal fallback. Raw score, not length adjusted. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| HealthBench (length adjusted) | 60% Section 8.17: same max-effort evaluation, length-adjusted score. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| HealthBench Professional (raw) | 74.2% Section 8.17: five trials, max effort, Opus 4.8 grader, safety classifiers and Opus 5 fallback. Raw score. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| HealthBench Professional (length adjusted) | 62.1% Section 8.17: same max-effort evaluation, length-adjusted score. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| BenchCAD Vision2Code (without tools) | 0.437voxel IoU Section 8.14.2: random 1,000-file subset, average five runs, adaptive thinking max effort, no tools. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| BenchCAD Vision2Code (with tools) | 0.843voxel IoU Section 8.14.2: random 1,000-file subset, average five runs, adaptive thinking max effort, container and image cropping tool. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| GDPval-AA v2 | 1853Elo Section 8.15.3: AA independent agentic professional work evaluation, max effort. Historical v2, not v2.1. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| AA-Briefcase (system card) | 1694Elo Section 8.15.4: independent AA long-horizon work evaluation, max effort; historical system-card snapshot. Not the medium-effort default catalog version. Anthropic Fable 5.1 system card · sections 8.1–8.11 | — |
| FrontierCode v1.1 (Extended) | 63.6% Cognition FrontierCode 1.1; Claude Fable 5.1; medium effort; claude-code harness; Extended subset (150 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.636. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget. Cognition · FrontierCode 1.1 original leaderboard | — |
| FrontierCode v1.1 (Main) | 50.91% Cognition FrontierCode 1.1; Claude Fable 5.1; medium effort; claude-code harness; Main subset (100 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.5091. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget. Cognition · FrontierCode 1.1 original leaderboard | — |
| Terminal-Bench 4.0 | 55.8% Section 8.6: Claude Code --bare, maximum thinking effort; 15 trials/task (990 trials), SE ±1.6–2 points. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| Terminal-Bench-Science 0.1 | 52.6% Section 8.7: Claude Code --bare, maximum thinking effort; 10 trials/task (700 trials), SE ±3.5–4.5 points. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| Humanity’s Last Exam (with tools) | 65% Table 8.1.A and section 8.12.1: web search, web fetch and code execution; thinking auto, 1M token cap, Opus 4.6 grader and source blocklist. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| Humanity’s Last Exam (without tools) | 60.9% Table 8.1.A: no tools; standard adaptive thinking at max effort, default sampling, five trials. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| AutomationBench | 31.4% Table 8.1.A: adaptive thinking at max effort, default sampling, five trials; production safeguards enabled. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| Chartography (with tools) | 86.2% Section 8.14.1: adaptive thinking at max effort; container and image-cropping tool, five runs. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| Chartography (without tools) | 42.6% Section 8.14.1: adaptive thinking at max effort, no tools, five runs. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| OSWorld 2.0 (partial) | 77.9% Section 8.14.3: August 2026 tasks with subsequent fixes; 1080p, maximum 500 action steps, maximum reasoning effort, five runs, Opus 4.8 grader. Not comparable to OSWorld 2.1 or earlier task releases. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
| OSWorld 2.0 (strict) | 41.7% Section 8.14.3: August 2026 tasks with subsequent fixes; 1080p, maximum 500 action steps, maximum reasoning effort, five runs, Opus 4.8 grader. Not comparable to OSWorld 2.1 or earlier task releases. Published Fable 5.1 evaluation, not the default medium catalog configuration or a TokenBazaar run; benchmark tools are not hosted API capabilities. Anthropic · Fable 5.1 system card | — |
No overall score, winner or cost-versus-score frontier is calculated without comparable, attributable benchmark evidence. Different test configurations appear separately. Vendor capabilities are distinct from TokenBaazar API support.
— means no published value or no matching capability. Provider token limits are not tested gateway guarantees; TokenBaazar media controls and response formats are listed separately.
Claude Fable 5.1
Sources and verification
Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.
- Provider model specifications
Vendor-reported
Last checked - Anthropic Fable 5.1 system card · sections 8.1–8.11
Vendor-reported
Last checked - CursorBench public evaluation
Third-party report
Last checked - Proximal · FrontierSWE public leaderboard
Third-party report
Last checked - Vals AI · Claude Fable 5.1 evaluation
Third-party report
Last checked - Google Gemini 4 Argon comparison chart
Vendor-reported
Last checked - Cognition · FrontierCode 1.1 original leaderboard
Third-party report
Last checked
Research notes (2)
- Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.
- Benchmark configuration is stated per result. Provider max-effort tests are not measurements of the default medium or xhigh catalog versions. ProgramBench 87.6 is the vendor’s filtered 166-task test, not a generic public score.