Models and Pricing

Compare models

Choose two or three models, including exact thinking and fast versions. Copy the URL to share your selection.

ComparisonGemini 3.1 Flash Lite
google/gemini-3.1-flash-lite
Choose model 2
TypeChat—
TokenBaazar price · 50% off₹12.00 in / ₹72.00 out per 1M tokens—
Market price₹24.00 in / ₹144.00 out per 1M tokens—
Context window (provider)Not applicable—
Input token limit (provider)1,048,576 tokens—
Maximum output65,536 tokens—
Vendor input modalitiesNot verified—
Vendor output modalitiesNot verified—
TokenBaazar endpointsPOST /v1/chat/completions—
Benchmark evidence3 source-linked results—
GPQA Diamond
86.9%

Official Google-reported score at launch (March 2026), thinking levels enabled per model defaults. Published source text: 86.9% on GPQA Diamond and 76.8% on MMMU Pro

Google Blog – Gemini 3.1 Flash-Lite: Built for intelligence at scale
—
MMMU Pro
76.8%

Official Google-reported score at launch (March 2026). Published source text: 76.8% on MMMU Pro

Google Blog – Gemini 3.1 Flash-Lite: Built for intelligence at scale
—
Arena.ai Elo
1432elo

Official Google-reported score at launch (March 2026). Published source text: 3.1 Flash-Lite achieves an impressive Elo score of 1432 on the Arena.ai Leaderboard

Google Blog – Gemini 3.1 Flash-Lite: Built for intelligence at scale
—

No overall score, winner or cost-versus-score frontier is calculated without comparable, attributable benchmark evidence. Different test configurations appear separately. Vendor capabilities are distinct from TokenBaazar API support.

— means no published value or no matching capability. Provider token limits are not tested gateway guarantees; TokenBaazar media controls and response formats are listed separately.

Gemini 3.1 Flash Lite

Sources and verification

Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.

Research notes (1)
  • Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.