← All models Model directory

AI model

Google: Gemini 3.7 Flash

google/gemini-3.7-flash

Google: Gemini 3.7 Flash is a 1,048,576-token model served on Composite at $0.120 per million tokens in and $1.88 per million tokens out. Ranked #34 of 203 models by traffic on Composite over the last 30 days, taking 0.43% of tokens and 0.26% of requests. Measured on Composite, 88.2% of its last 17 requests on Composite succeeded, the typical first token arrives in 28812 ms, output runs at about 73.0 tokens per second.

Try in Playground

Community rating

— No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDgoogle/gemini-3.7-flash
ProviderGoogle
Context window1,048,576 tokens
Acceptstext, image, video, file, audio
Returnstext
Input price$0.120 per million tokens
Output price$1.88 per million tokens
Prompt cachingSupported — cached input is 60% off the list rate on pay-as-you-go credits
Image input$3.8e-7 per image
Audio input$3.8e-7 per second
Reasoning tokens$1.88 per million tokens
Web search$0.0140 per result
ThinkingAlways on for this model
ModeratedNo
Cost on a monthly plan5 requests of the daily allowance per call
Plans that may spend on itBasic, Plus, Pro, Ultra, Enterprise
Routing endpoints2 (base plus 1 routing variant)

Traffic rank on Composite

#34 of 203

0.43% of tokens and 0.26% of requests over 30 days.

Success rate

88.2%

Across the last 17 requests Composite served for this model.

Time to first token

28812 ms

Median across the same sample.

Output speed

73.0 tok/s

Median generation rate after the first token.

Routing variants

The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
cheapThink google/gemini-3.7-flash:cheapThink Google $0.120 per million tokens input · $1.88 per million tokens output

Pricing

$0.120 per million tokens input · $1.88 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 5 requests of that day's allowance. That needs the Basic or Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.

Provider

Google

Roleplay fit

Gemini models combine large context windows with fast throughput, which keeps response times low even in busy group chats.

Context behavior

With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.

From the provider

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Supplied by Google, not written by Composite.