Models and Pricing

Compare models

Choose two or three models, including exact thinking and fast versions. Copy the URL to share your selection.

ComparisonClaude Haiku 4.5
anthropic/claude-haiku-4-5
Choose model 2
TypeChat—
TokenBaazar price · 50% off₹48.00 in / ₹240.00 out per 1M tokens—
Market price₹96.00 in / ₹480.00 out per 1M tokens—
Context window (provider)200,000 tokens—
Input token limit (provider)Within the context window—
Maximum output64,000 tokens—
Knowledge cutoffFeb 2025—
Vendor input modalitiesText, Image—
Vendor output modalitiesText—
TokenBaazar endpointsPOST /v1/chat/completions · POST /v1/messages—
TokenBaazar inputText, Image/PDF content through the chat bridge—
TokenBaazar outputText—
Benchmark evidence1 source-linked results—
SWE-bench Verified
73.3%

Anthropic methodology footnote: simple bash+file-edit scaffold, averaged over 50 trials, no test-time compute, 128K thinking budget, default sampling, full 500-problem set, with a prompt addendum instructing heavy tool use. Haiku 4.5's catalog default is Extended thinking (no explicit low/medium/high effort knob), so this is treated as the base/default-thinking configuration, not a stripped-down no-thinking run.

Anthropic · Introducing Claude Haiku 4.5
—

No overall score, winner or cost-versus-score frontier is calculated without comparable, attributable benchmark evidence. Different test configurations appear separately. Vendor capabilities are distinct from TokenBaazar API support.

— means no published value or no matching capability. Provider token limits are not tested gateway guarantees; TokenBaazar media controls and response formats are listed separately.

Claude Haiku 4.5

Sources and verification

Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.

Research notes (1)
  • Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.