Compare models
Choose two or three models, including exact thinking and fast versions. Copy the URL to share your selection.
| Comparison | GPT-6 Astra Fastopenai/gpt-6-astra-fast | Choose model 2 |
|---|---|---|
| Type | Chat | — |
| Best for | Fast, low-cost replies and high-volume everyday tasks | — |
| TokenBazaar price | $20.00 in / $100.00 out per 1M tokens | — |
| Cached input / 1M tokens | $2.00 per 1M tokens | — |
| Cache write / 1M tokens | $25.00 per 1M tokens | — |
| Cached output | Not offered · standard output rate applies | — |
| Market price | $20.00 in / $100.00 out per 1M tokens | — |
| Context window (provider) | 1,050,000 tokens | — |
| Input token limit (provider) | Within the context window | — |
| Maximum output | 128,000 tokens | — |
| Knowledge cutoff | Apr 30, 2026 | — |
| Vendor input modalities | Text, Image | — |
| Vendor output modalities | Text | — |
| TokenBazaar endpoints | POST /v1/chat/completions | — |
| TokenBazaar input | Text, Image/PDF content through the chat bridge | — |
| TokenBazaar output | Text | — |
No overall score, winner or cost-versus-score frontier is calculated without comparable, attributable benchmark evidence. Different test configurations appear separately. Vendor capabilities are distinct from TokenBazaar API support.
— means no published value or no matching capability. Provider token limits are not tested gateway guarantees; TokenBazaar media controls and response formats are listed separately.
GPT-6 Astra Fast
Sources and verification
Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.
- Provider model specifications
Vendor-reported
Last checked - Vellum · GPT-6 Astra launch-table analysis
Third-party report
Last checked - ARC Prize · GPT-6 Astra evaluated results
Third-party report
Last checked - Cognition · FrontierCode 1.1 original leaderboard
Third-party report
Last checked
Research notes (1)
- Published provider limits; maximum-capacity requests have not been independently exercised through TokenBazaar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.