Traffic rank on Composite
#119 of 203
0.01% of tokens and 0% of requests over 30 days.
AI model
nvidia/nemotron-3.5-lightning
NVIDIA: Nemotron 3.5 Lightning is a 262,144-token model served on Composite at $0.0157 per million tokens in and $0.140 per million tokens out. Ranked #119 of 203 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0% of requests. Measured on Composite, 100.0% of its last 3 requests on Composite succeeded, the typical first token arrives in 49187 ms, output runs at about 280.9 tokens per second.
Try in Playground| Model ID | nvidia/nemotron-3.5-lightning |
|---|---|
| Provider | cv11 |
| Context window | 262,144 tokens |
| Accepts | text |
| Returns | text |
| Input price | $0.0157 per million tokens |
| Output price | $0.140 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Moderated | No |
| Cost on a monthly plan | 1 request of the daily allowance per call |
| Plans that may spend on it | Every plan, plus pay-as-you-go credits |
| Routing endpoints | 10 (base plus 9 routing variants) |
#119 of 203
0.01% of tokens and 0% of requests over 30 days.
100.0%
Across the last 3 requests Composite served for this model.
49187 ms
Median across the same sample.
280.9 tok/s
Median generation rate after the first token.
The same model behind 9 alternative endpoints. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| Normal | nvidia/nemotron-3.5-lightning-normal |
cv11 | $0.0192 per million tokens input · $0.160 per million tokens output |
| Normal + cheapThink | nvidia/nemotron-3.5-lightning-normal:cheapThink |
cv11 | $0.0192 per million tokens input · $0.160 per million tokens output |
| Cheapest provider | nvidia/nemotron-3.5-lightning:cheapest-provider |
Io Net | $0.0157 per million tokens input · $0.140 per million tokens output |
| Cheapest provider + cheapThink | nvidia/nemotron-3.5-lightning:cheapest-provider:cheapThink |
Io Net | $0.0157 per million tokens input · $0.140 per million tokens output |
| cheapThink | nvidia/nemotron-3.5-lightning:cheapThink |
cv11 | $0.0157 per million tokens input · $0.140 per million tokens output |
| Fast | nvidia/nemotron-3.5-lightning:fast |
Darkbloom | $0.0125 per million tokens input · $0.180 per million tokens output |
| Fast + cheapThink | nvidia/nemotron-3.5-lightning:fast:cheapThink |
Darkbloom | $0.0125 per million tokens input · $0.180 per million tokens output |
| Quality | nvidia/nemotron-3.5-lightning:quality |
CoreWeave | $0.0224 per million tokens input · $0.200 per million tokens output |
| Quality + cheapThink | nvidia/nemotron-3.5-lightning:quality:cheapThink |
CoreWeave | $0.0224 per million tokens input · $0.200 per million tokens output |
$0.0157 per million tokens input · $0.140 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 1 request of that day's allowance.
cv11
Nemotron models are tuned by NVIDIA for instruction-following and tool use, and tend to behave predictably in structured prompts.
With a 262,144-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Supplied by cv11, not written by Composite.
NVIDIA: Nemotron 3.5 Content Safety · NVIDIA: Nemotron 3 Ultra · NVIDIA: Nemotron 3 Super