OpenRating is an independent, open-source AI rating platform. Join the community

OOpenRating
All models

Google · Released Feb 25, 2026

Gemini 3.7 Flash (high)

Google's blazing-fast breakthrough model delivering near-400 tok/s output speed with 1M context and frontier reasoning benchmarks.

Proprietary

OpenRating Index

56.2

0–100 composite

Reasoning Index

75.8

0–100 composite

Cost / Task

$0.40

benchmark run

Output speed

392 tok/s

blended tok/s

Time to first token

0.4s

median latency

Input price

$0.3

per 1M tokens

Output price

$1.2

per 1M tokens

Context window

1M

tokens

Strengths

  • Industry-leading 391+ tok/s output speed

  • 1,000,000 token context window for huge video & doc corpora

  • Ultra-low task cost at $0.40 per benchmark run

How we rate Gemini 3.7 Flash (high)

OpenRating scores every model on the same axes. The OpenRating Index is a weighted composite of reasoning, knowledge, and agentic benchmarks, normalised to a 0–100 scale. Speed is blended output measured in tokens per second under standardised load, with time-to-first-token reported separately. Prices are list input and output prices per million tokens, and Task Cost measures end-to-end workload spending.

OpenRating Index

56.2

Reasoning Index

75.8

Token Pricing

Input per 1M tokens

$0.3

Output per 1M tokens

$1.2

Cost of 1M input + 1M output

$1.5

Task Cost Economics

Empirical cost breakdown for a standard OpenRating Index benchmark task ($0.40 total).

Reasoning (Chain of Thought)
$0.053
Answer generation
$0.086
Cache write overhead
$0.16
Cache hit queries
$0.098
Input payload
$0.005