Models and Pricing

Compare models

Choose two or three models, including exact thinking and fast versions. Copy the URL to share your selection.

ComparisonClaude Opus 5.5 Max
anthropic/claude-opus-5-5-xhigh
Choose model 2
TypeChat—
TokenBaazar price · 50% off₹192.00 in / ₹960.00 out per 1M tokens—
Market price₹384.00 in / ₹1920.00 out per 1M tokens—
Context window (provider)1,000,000 tokens—
Input token limit (provider)Within the context window—
Maximum output128,000 tokens—
Knowledge cutoffJun 2026—
Vendor input modalitiesText, Image—
Vendor output modalitiesText—
TokenBaazar endpointsPOST /v1/chat/completions · POST /v1/messages—
TokenBaazar inputText, Image/PDF content through the chat bridge—
TokenBaazar outputText—
Benchmark evidence5 source-linked results—
Terminal-Bench 4.0
60%

Explicitly the Xhigh effort configuration ('Claude Opus 5.5 (Xhigh, Default Fallback)') on AA's independently-run Terminal-Bench 4.0 harness — matches the requested Opus 5.5 xhigh condition exactly, not max/fast/default. Distinct from Anthropic's own self-reported xhigh figure of 66.4% already on file (different methodology/harness) and from AA's narrative max-effort figure of 59.6% quoted in AA's launch article.

Artificial Analysis · Opus 5.5 (Xhigh) vs Opus 5 (High) comparison
—
Terminal-Bench-Science 0.1
62%

Explicitly the xhigh effort configuration ('Claude Opus 5.5 (xhigh)'); AA also separately reports 59% at max effort for the same model on this benchmark — kept distinct and not merged. Independent AA harness, not Anthropic's own 58.7% (max effort) figure already on file.

Artificial Analysis · Terminal-Bench-Science 0.1 leaderboard launch post (LinkedIn)
—
CursorBench 4.0
56%

Row explicitly labeled 'Opus 5.5 Extra High' (xhigh) at 56.0%, $6.98/task. Kept distinct from the existing 57.8% entry, which is Anthropic's own self-reported max-effort figure from the launch page, and from CursorBench's own 'Opus 5.5 Max' row (57.8%, $13.43/task) — same number as Anthropic's but a separately sourced/priced data point.

Cursor · CursorBench 4.0 public leaderboard
—
FrontierCode v1.1 (Main)
51.42%

Cognition FrontierCode 1.1; Claude Opus 5.5; xhigh effort; claude-code harness; Main subset (100 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.5142. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget.

Cognition · FrontierCode 1.1 original leaderboard
—
FrontierCode v1.1 (Extended)
63.53%

Cognition FrontierCode 1.1; Claude Opus 5.5; xhigh effort; claude-code harness; Extended subset (150 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.6353. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget.

Cognition · FrontierCode 1.1 original leaderboard
—

No overall score, winner or cost-versus-score frontier is calculated without comparable, attributable benchmark evidence. Different test configurations appear separately. Vendor capabilities are distinct from TokenBaazar API support.

— means no published value or no matching capability. Provider token limits are not tested gateway guarantees; TokenBaazar media controls and response formats are listed separately.

Claude Opus 5.5 Max

Sources and verification

Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.

Research notes (1)
  • Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.