Compare model

google / Chat

Gemini 3.7 Flash

google/gemini-3.7-flash

Latest Gemini Flash (default)

Catalog status

Enabled

Research checked

2026-10-07

Price / performance

What it costs to get this result

FrontierCode v1.1 (Extended) · Source-reported results, not a composite ranking. Prices use the average of TokenBazaar input and output rates per million tokens; actual spend depends on usage. Test configurations differ.

Gemini 3.7 Flash56.27% · 43 sourced models

Blended price / 1M tokens · log scale

● Current model● Other modelsDashed line: price/score frontier among plotted models
Compare sourced models on this benchmark

Benchmark results

7 reported results

Results are source-attributed, not TokenBazaar measurements. Different tests and thinking variants are not interchangeable.

7 of 7 results

FrontierCode 1.1 Main

Coding · Vendor-reported

Test conditions and source

Official Google-reported score at launch, Aug 13 2026, vs Gemini 3.6 Flash baseline of 34.4%. Published source text: improved performance in generating production-ready code as seen in FrontierCode 1.1 Main (43.6% vs 34.4%)

Tested model: google/gemini-3.7-flash

Google Blog – Introducing Gemini 3.7 Flash · Checked 2026-10-06

43.6%

DeepSWE v1.1

Coding · Vendor-reported

Test conditions and source

Official Google-reported score at launch, Aug 13 2026; same DeepSWE v1.1 effort/harness used for Gemini 3.6 Flash's 49.0% baseline. Published source text: DeepSWE v1.1 (65.3% vs 49.0%)

Tested model: google/gemini-3.7-flash

Google Blog – Introducing Gemini 3.7 Flash · Checked 2026-10-06

65.3%

WebDev Arena Elo

Coding · Vendor-reported

Test conditions and source

Official Google-reported score at launch, Aug 13 2026. Published source text: outperforms 3.6 Flash on Arena.ai's WebDev Arena with an Elo score of 1588 vs 1538

Tested model: google/gemini-3.7-flash

Google Blog – Introducing Gemini 3.7 Flash · Checked 2026-10-06

1588elo

GDP.pdf

Knowledge · Vendor-reported

Test conditions and source

Official Google-reported score at launch, Aug 13 2026. Published source text: significantly outperforms 3.6 Flash on the GDP.pdf benchmark (34.0% vs 22.0%)

Tested model: google/gemini-3.7-flash

Google Blog – Introducing Gemini 3.7 Flash · Checked 2026-10-06

34%

AutomationBench

Agentic · Vendor-reported

Test conditions and source

Official Google-reported score at launch, Aug 13 2026. Published source text: surpasses 3.6 Flash in AutomationBench ... (30.4% vs 17.0%)

Tested model: google/gemini-3.7-flash

Google Blog – Introducing Gemini 3.7 Flash · Checked 2026-10-06

30.4%

FrontierCode v1.1 (Extended)

Coding · Third-party report

Test conditions and source

Cognition FrontierCode 1.1; Gemini 3.7 Flash; medium effort; chisel harness; Extended subset (150 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.5627. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget.

Tested model: google/gemini-3.7-flash

Cognition · FrontierCode 1.1 original leaderboard · Checked 2026-10-07

56.27%

FrontierCode v1.1 (Main)

Coding · Third-party report

Test conditions and source

Cognition FrontierCode 1.1; Gemini 3.7 Flash; medium effort; chisel harness; Main subset (100 tasks). Published weighted rubric score (new_score × 100), not the all-or-nothing pass rate; solution-source violations score zero. Source field new_score=0.4359. Publisher evaluation, not a TokenBazaar measurement; provider effort configuration is not a guarantee of the API's fixed thinking budget.

Tested model: google/gemini-3.7-flash

Cognition · FrontierCode 1.1 original leaderboard · Checked 2026-10-07

43.59%

Model specifications

Input token limit

1.05M

Published provider limit

Maximum output

65.54K

Published provider limit

Input and output modalities

Provider-reported formats; API compatibility is detailed below.

Input

Not verified

Output

Not verified
Exact model ID
google/gemini-3.7-flash
Model type
Chat
Input token limit
1,048,576 tokens
Maximum output
65,536 tokens

Input token limits describe provider input capacity. Published provider limits are not independently tested TokenBazaar request limits.

Versions and thinking levels

One enabled version is currently listed for this model.

Gemini 3.7 FlashSelected₹36.00 in / ₹180.00 out per 1M tokens

Features and tool support

TokenBazaar’s exposed interface, not every feature advertised by the provider. “Available” describes the implemented interface, not a successful test of every input or tool.

Reasoning configuration

No independent reasoning-mode guarantee is recorded for this exact model.

See enabled versions and provider sources for supported controls.

Coding and agent workflows

Text/code generation and customer-managed tool loops.

Coding benchmark results do not establish a hosted terminal, sandbox, autonomous browser or guaranteed task success.

Streaming answers

Available

Incremental text through TokenBazaar’s chat endpoint; Claude also has native Messages access.

Customer-defined function tools

Limited

Only customer-defined function tools are forwarded. Model/protocol compatibility applies; your application authorizes and executes every tool.

Forced or named tool selection

Limited

The chat route accepts auto, none and required. A named-tool choice is not forwarded; exact model compatibility is not independently tested.

Images and PDFs

Not verified

The chat bridge accepts image/PDF content. Provider input modalities are listed separately; file size, content and model limits still apply.

Hosted web search / browsing

Not exposed

No provider-hosted web search or web-fetch tool is exposed. Customer-owned search can be implemented as an authorized function tool.

Hosted code execution / computer use

Not exposed

No built-in execution environment or computer-control tool is provided. Code generation is not code execution.

Hosted file search / persistent assistants

Not exposed

No hosted file-search index, Assistants endpoint or persistent provider agent is exposed.

Batch / fine-tuning / cached-price discounts

Not exposed

No public batch or fine-tuning endpoint, or separate cached-token discount, is offered here.

Sources and verification

Provider specifications describe the upstream model, not a guarantee of every feature through TokenBazaar. Prices come from the enabled catalog; benchmark results belong to the exact tested model.

Research notes (1)
  • Published provider limits; maximum-capacity requests have not been independently exercised through TokenBaazar. Thinking-level variants use the same provider model; reasoning and answer tokens share the output budget.