Traffic rank on Composite
#113 of 203
0.01% of tokens and 0.03% of requests over 30 days.
AI model
google/gemma-4-31b-it:free
Google: Gemma 4 31B (free) is a 262,144-token model free on Composite. Ranked #113 of 203 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0.03% of requests. 1 up and 2 down from 3 votes. A positive score is published once a model reaches 5 votes. Measured on Composite, 47.0% of its last 100 requests on Composite succeeded, the typical first token arrives in 2628 ms, output runs at about 52.3 tokens per second.
Try in Playground| Model ID | google/gemma-4-31b-it:free |
|---|---|
| Provider | cv11 |
| Context window | 262,144 tokens |
| Accepts | image, text, video |
| Returns | text |
| Input price | Free |
| Output price | Free |
| Prompt caching | Not offered for this model |
| Moderated | No |
| Cost on a monthly plan | Free — does not spend the daily request allowance |
| Plans that may spend on it | Every plan, plus pay-as-you-go credits |
| Routing endpoints | 1 |
#113 of 203
0.01% of tokens and 0.03% of requests over 30 days.
47.0%
Across the last 100 requests Composite served for this model.
2628 ms
Median across the same sample.
52.3 tok/s
Median generation rate after the first token.
0.0%
Share of prompt tokens served from the provider's cache over 7 days.
Free model
Input and output are charged at the provider’s list price.
Monthly plans serve this model without spending any of their daily requests.
cv11
Gemini models combine large context windows with fast throughput, which keeps response times low even in busy group chats.
With a 262,144-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It costs nothing to use on Composite's free tier, which makes it a reasonable place to test a character card or system prompt before spending on a paid model.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Supplied by cv11, not written by Composite.
Composite Gemini Router · Google: Gemini 3.8 Flash (Cheap) · Google: Gemini 3.7 Flash (Cheap)