Traffic rank on Composite
#10 of 203
3.37% of tokens and 3.75% of requests over 30 days.
AI model
nvidia/nemotron-3-super-120b-a12b:free
NVIDIA: Nemotron 3 Super (free) is a 262,144-token model free on Composite. Ranked #10 of 203 models by traffic on Composite over the last 30 days, taking 3.37% of tokens and 3.75% of requests. 2 up and 1 down from 3 votes. A positive score is published once a model reaches 5 votes. Measured on Composite, 100.0% of its last 100 requests on Composite succeeded, the typical first token arrives in 12272 ms, output runs at about 200.1 tokens per second.
Try in Playground| Model ID | nvidia/nemotron-3-super-120b-a12b:free |
|---|---|
| Provider | cv11 |
| Context window | 262,144 tokens |
| Accepts | text |
| Returns | text |
| Input price | Free |
| Output price | Free |
| Prompt caching | Not offered for this model |
| Thinking | Optional - see "NVIDIA: Nemotron 3 Super (free) (Reasoning)" in the catalog to turn it on by default |
| Moderated | No |
| Cost on a monthly plan | Free — does not spend the daily request allowance |
| Plans that may spend on it | Every plan, plus pay-as-you-go credits |
| Routing endpoints | 1 |
#10 of 203
3.37% of tokens and 3.75% of requests over 30 days.
100.0%
Across the last 100 requests Composite served for this model.
12272 ms
Median across the same sample.
200.1 tok/s
Median generation rate after the first token.
25.8%
Share of prompt tokens served from the provider's cache over 7 days.
Free model
Input and output are charged at the provider’s list price.
Monthly plans serve this model without spending any of their daily requests.
cv11
Nemotron models are tuned by NVIDIA for instruction-following and tool use, and tend to behave predictably in structured prompts.
With a 262,144-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It costs nothing to use on Composite's free tier, which makes it a reasonable place to test a character card or system prompt before spending on a paid model.
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Supplied by cv11, not written by Composite.
NVIDIA: Nemotron 3.5 Lightning · NVIDIA: Nemotron 3.5 Content Safety · NVIDIA: Nemotron 3 Ultra