Traffic rank on Composite
#34 of 203
0.43% of tokens and 0.26% of requests over 30 days.
AI model
google/gemini-3.7-flash
Google: Gemini 3.7 Flash is a 1,048,576-token model served on Composite at $0.120 per million tokens in and $1.88 per million tokens out. Ranked #34 of 203 models by traffic on Composite over the last 30 days, taking 0.43% of tokens and 0.26% of requests. Measured on Composite, 88.2% of its last 17 requests on Composite succeeded, the typical first token arrives in 28812 ms, output runs at about 73.0 tokens per second.
Try in Playground| Model ID | google/gemini-3.7-flash |
|---|---|
| Provider | |
| Context window | 1,048,576 tokens |
| Accepts | text, image, video, file, audio |
| Returns | text |
| Input price | $0.120 per million tokens |
| Output price | $1.88 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Image input | $3.8e-7 per image |
| Audio input | $3.8e-7 per second |
| Reasoning tokens | $1.88 per million tokens |
| Web search | $0.0140 per result |
| Thinking | Always on for this model |
| Moderated | No |
| Cost on a monthly plan | 5 requests of the daily allowance per call |
| Plans that may spend on it | Basic, Plus, Pro, Ultra, Enterprise |
| Routing endpoints | 2 (base plus 1 routing variant) |
#34 of 203
0.43% of tokens and 0.26% of requests over 30 days.
88.2%
Across the last 17 requests Composite served for this model.
28812 ms
Median across the same sample.
73.0 tok/s
Median generation rate after the first token.
The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| cheapThink | google/gemini-3.7-flash:cheapThink |
$0.120 per million tokens input · $1.88 per million tokens output |
$0.120 per million tokens input · $1.88 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 5 requests of that day's allowance. That needs the Basic or Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.
Gemini models combine large context windows with fast throughput, which keeps response times low even in busy group chats.
With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Supplied by Google, not written by Composite.
Composite Gemini Router · Google: Gemini 3.8 Flash (Cheap) · Google: Gemini 3.7 Flash (Cheap)