nvidia
NVIDIA: Nemotron 3.5 Lightning
nvidia/nemotron-3.5-lightning
For $1, you can send approximately:
~408messages
How do we get this number?
One message = ~7,000 input tokens + ~7,000 output tokens
Input cost per message7,000 x $0.10/M = $0.000700
Output cost per message7,000 x $0.25/M = $0.001750
Total cost per message$0.002450
Messages for $1408.16
Context window
262k
tokens
Max response
236k
tokens
Input price
$0.10
per million tokens
Output price
$0.25
per million tokens
Modalities
Input:textOutput:text
Description
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...