Nvidia: Nemotron 3.5 Lightning (Free)
Nvidia: Nemotron 3.5 Lightning (Free) is a LLM model.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that... **Terms of service** For NVIDIA free endpoints (Super/Ultra/etc): Trial use only - do not submit personal or confidential data. Your use is logged for security purposes and to improve NVIDIA products and services. The logged session data for improvement purposes is not linked to your identity or any persistent identifier. For more information about our data processing practices, see our [Privacy Policy](https://www.nvidia.com/en-us/about-nvidia/privacy-policy/). By interacting with this endpoint, you consent to our collection, recording, and use of such information and the [NVIDIA API Trial Terms of Service](https://assets.ngc.nvidia.com/products/api-catalog/legal/NVIDIA%20API%20Trial%20Terms%20of%20Service.pdf).
Vision
Audio
Tool Use
Reasoning
Citations
Key Specs
Context
1M
Max Output
262k
Best Price (Input / Output)
Chat with Model
$0.05 / $0.20 per 1M tokens
Pricing Comparison
| Provider | Input (1M) | Output (1M) | Image (1k) |
|---|---|---|---|
| kilo (Free Tier) | Free | Free | - |
| opencode (Free Tier) | Free | Free | - |
| openrouter (Free Tier) | Free | Free | - |
| nano-gpt | $0.05 | $0.20 | - |
| openrouter | $0.10 | $0.25 | - |
| wandb | $0.10 | $0.25 | - |
Prices are per 1 million tokens unless otherwise noted. Image pricing is per 1000 images if applicable.
Price History
Prompt Price
Completion Price
Price trend per 1 million tokens over time.
Supported Parameters
frequency_penalty
reasoning
min_p
structured_outputs
response_format
include_reasoning
stop
top_k
max_tokens
top_logprobs
temperature
presence_penalty
tools
seed
tool_choice
logit_bias
logprobs
repetition_penalty
top_p