← All Models

NVIDIA: Nemotron 3.5 Lightning

🇺🇸 NVIDIA · Nemotron 3.5

Input Price $0.080 per million tokens NT$2.6
Output Price $0.200 per million tokens NT$6.4
Context Window 262K tokens Output limit: 131K
OpenRouter Route Price Please verify with official pricing pages
Use this model via OpenRouter →

Overview

NVIDIA: Nemotron 3.5 Lightning is a large language model API from NVIDIA, part of its Nemotron 3.5 model family. At $0.080 per million input tokens and $0.200 per million output tokens, it sits in the budget tier — among the cheaper options for high-throughput or cost-sensitive workloads. Output tokens cost about 3× as much as input, so prompt-heavy workloads run noticeably cheaper than generation-heavy ones. A large 262K-token context window (≈393 pages of text) lets it take in whole books, large codebases, or lengthy transcripts in a single call. On Artificial Analysis's Intelligence Index it scores 24 (D grade), a useful proxy for its general reasoning strength relative to the other models tracked here. All prices on this page reflect OpenRouter's routed rates and are re-synced automatically every day; confirm against the provider's official pricing before committing to production.

Dimension Unit Price (USD) Price (TWD) Effective From
Input per 1M tokens $0.080 NT$2.6
Output per 1M tokens $0.200 NT$6.4
Cached Input per 1M tokens $0.040 NT$1.3

Provider
NVIDIA
Model Family
Nemotron 3.5
Version String
nvidia/nemotron-3.5-lightning
Status
Active
Modality
Text
Context Window
262,144 tokens
Output Limit
131,072 tokens

Index Metrics

Cross-domain capability indexes evaluated by Artificial Analysis — Artificial Analysis

Agentic Index 14 F Measured: 2026-08-26
Coding Index 27 C Measured: 2026-08-26
Intelligence Index 24 D Measured: 2026-08-26

Benchmark Scores

Data source: Artificial Analysis

AA-LCR 55.3% C Measured: 2026-08-26
GPQA Diamond 74.3% B Measured: 2026-08-26
Non-Hallucination 62.4% Measured: 2026-08-26
Omniscience Accuracy 14.4% Measured: 2026-08-26
SciCode 31.6% B Measured: 2026-08-26

Performance Metrics

Real-world benchmarks, updated every 72 hours by Artificial Analysis — Artificial Analysis

First Token Latency 1.1s Measured: 2026-08-26
Output Speed 305 t/s Measured: 2026-08-26
Response Time 9.3s Measured: 2026-08-26

90-Day Price Trend

Input / Output price (USD per 1M tokens)

Past 90 days of records; every price change is shown here

Date Dimension Price (USD) Source
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.040 OpenRouter
Output $0.200 OpenRouter
Input $0.080 OpenRouter
Cached Input $0.050 OpenRouter
Output $0.250 OpenRouter
Input $0.100 OpenRouter
Cached Input $0.050 OpenRouter
Output $0.250 OpenRouter
Input $0.100 OpenRouter
Cached Input $0.050 OpenRouter
Output $0.250 OpenRouter
Input $0.100 OpenRouter
Cached Input $0.050 OpenRouter
Output $0.250 OpenRouter
Input $0.100 OpenRouter
Cached Input $0.050 OpenRouter
Output $0.250 OpenRouter
Input $0.100 OpenRouter

Key Insights

Key data points from this page for quick reference and citation.

  • NVIDIA: Nemotron 3.5 Lightning Input price: $0.08/M tokens
  • NVIDIA: Nemotron 3.5 Lightning Output price: $0.2/M tokens
  • Context window: 262,144 tokens
  • Provider: NVIDIA
  • Model family: Nemotron 3.5
  • Modalities: Text
  • Data source: OpenRouter, updated daily