← All models Model directory

AI model

Google: Gemma 4 31B (free)

google/gemma-4-31b-it:free

Google: Gemma 4 31B (free) is a 262,144-token model free on Composite. Ranked #113 of 203 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0.03% of requests. 1 up and 2 down from 3 votes. A positive score is published once a model reaches 5 votes. Measured on Composite, 47.0% of its last 100 requests on Composite succeeded, the typical first token arrives in 2628 ms, output runs at about 52.3 tokens per second.

Try in Playground

Community rating

— 1 up and 2 down from 3 votes. A positive score is published once a model reaches 5 votes.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDgoogle/gemma-4-31b-it:free
Providercv11
Context window262,144 tokens
Acceptsimage, text, video
Returnstext
Input priceFree
Output priceFree
Prompt cachingNot offered for this model
ModeratedNo
Cost on a monthly planFree — does not spend the daily request allowance
Plans that may spend on itEvery plan, plus pay-as-you-go credits
Routing endpoints1

Traffic rank on Composite

#113 of 203

0.01% of tokens and 0.03% of requests over 30 days.

Success rate

47.0%

Across the last 100 requests Composite served for this model.

Time to first token

2628 ms

Median across the same sample.

Output speed

52.3 tok/s

Median generation rate after the first token.

Prompt cache hits

0.0%

Share of prompt tokens served from the provider's cache over 7 days.

Pricing

Free model

Input and output are charged at the provider’s list price.

Monthly plans serve this model without spending any of their daily requests.

Provider

cv11

Roleplay fit

Gemini models combine large context windows with fast throughput, which keeps response times low even in busy group chats.

Context behavior

With a 262,144-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It costs nothing to use on Composite's free tier, which makes it a reasonable place to test a character card or system prompt before spending on a paid model.

From the provider

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Supplied by cv11, not written by Composite.