Back to comparator
nvidia

NVIDIA: Nemotron 3.5 Lightning

nvidia/nemotron-3.5-lightning

For $1, you can send approximately:
~408messages

How do we get this number?

One message = ~7,000 input tokens + ~7,000 output tokens
Input cost per message7,000 x $0.10/M = $0.000700
Output cost per message7,000 x $0.25/M = $0.001750
Total cost per message$0.002450
Messages for $1408.16
Context window
262k
tokens
Max response
236k
tokens
Input price
$0.10
per million tokens
Output price
$0.25
per million tokens

Modalities

Input:textOutput:text

Description

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...