openai / Chat
GPT-5.5 Pro
openai/gpt-5.5-proExtended-reasoning GPT-5.5 Pro
Catalog status
Enabled
Research checked
2026-10-06
Price / performance
What it costs to get this result
ARC-AGI-2 · Source-reported results, not a composite ranking. Prices use the average of TokenBazaar input and output rates per million tokens; actual spend depends on usage. Test configurations differ.
Blended price / 1M tokens · log scale
Compare sourced models on this benchmark
Benchmark results
Results are source-attributed, not TokenBazaar measurements. Different tests and thinking variants are not interchangeable.
9 of 9 results
ARC-AGI-1
Reasoning · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. ARC-AGI-1 96.5% 92nd percentile · rank 9 of 97 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
ARC-AGI-2
Reasoning · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. ARC-AGI-2 84.6% 91st percentile · rank 10 of 99 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
BrowseComp
Agentic · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. BrowseComp 90.1% 88th percentile · rank 8 of 60 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
Humanity's Last Exam (with tools)
Reasoning · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "tools": true, "third_party_aggregator": true}. Humanity's Last Exam 57.2% with tools / 43.1% no tools Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
Humanity's Last Exam (no tools)
Reasoning · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "tools": false, "third_party_aggregator": true}. Humanity's Last Exam 57.2% with tools / 43.1% no tools Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
FrontierMath
Math · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. FrontierMath 2025-02-28 Private 52.4% 83rd percentile · rank 7 of 37 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
FrontierMath Tier 4
Math · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. FrontierMath Tier 4 2025-07-01 Private 39.6% 78th percentile · rank 6 of 24 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
GDPval
Knowledge-Work · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. GDPval 82.3% Generalization · 88th percentile · rank 3 of 18 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
SimpleBench
Reasoning · Third-party report
Test conditions and source
Published configuration: {"variant": "Pro", "third_party_aggregator": true}. SimpleBench 76.9% 94th percentile · rank 3 of 36 Published model-family result, not a test of TokenBazaar or its default effort.
Tested model: openai/gpt-5.5-pro
BenchmarkList - GPT-5.5 Pro · Checked 2026-10-06
Model specifications
Context window
1.05M
Published provider limit
Maximum output
128K
Published provider limit
Knowledge cutoff
Dec 01, 2025
Upstream model
Input and output modalities
Provider-reported formats; API compatibility is detailed below.
Input
Output
- Exact model ID
- openai/gpt-5.5-pro
- Model type
- Chat
- Context window
- 1,050,000 tokens
- Maximum output
- 128,000 tokens
- Knowledge cutoff
- Dec 01, 2025
A context window is the total conversation budget, not a separate maximum input allowance; generated output and reasoning can use that budget. Published provider limits are not independently tested TokenBazaar request limits.
Versions and thinking levels
Choose from 3 enabled versions. The exact ID determines the version and its price; benchmark scores do not carry over between variants.
Features and tool support
TokenBazaar’s exposed interface, not every feature advertised by the provider. “Available” describes the implemented interface, not a successful test of every input or tool.
Reasoning configuration
The catalog version determines the requested reasoning effort.
Selected effort: medium (default catalog version). Reasoning level is not a measured intelligence or speed score.
Coding and agent workflows
Text/code generation and customer-managed tool loops.
Coding benchmark results do not establish a hosted terminal, sandbox, autonomous browser or guaranteed task success.
Streaming answers
AvailableIncremental text through TokenBazaar’s chat endpoint; Claude also has native Messages access.
Customer-defined function tools
LimitedOnly customer-defined function tools are forwarded. Model/protocol compatibility applies; your application authorizes and executes every tool.
Forced or named tool selection
LimitedThe chat route accepts auto, none and required. A named-tool choice is not forwarded; exact model compatibility is not independently tested.
Images and PDFs
Not verifiedThe chat bridge accepts image/PDF content. Provider input modalities are listed separately; file size, content and model limits still apply.
Hosted web search / browsing
Not exposedNo provider-hosted web search or web-fetch tool is exposed. Customer-owned search can be implemented as an authorized function tool.
Hosted code execution / computer use
Not exposedNo built-in execution environment or computer-control tool is provided. Code generation is not code execution.
Hosted file search / persistent assistants
Not exposedNo hosted file-search index, Assistants endpoint or persistent provider agent is exposed.
Batch / fine-tuning / cached-price discounts
Not exposedNo public batch or fine-tuning endpoint, or separate cached-token discount, is offered here.
Sources and verification
Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.
- Provider model specifications
Vendor-reported
Last checked - BenchmarkList - GPT-5.5 Pro
Third-party report
Last checked
Research notes (1)
- Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.