Models, measured side by side
Grok 4.7 vs Nemotron 3.5 Lightning
Compare Grok 4.7, Nemotron 3.5 Lightning by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
Grok 4.7
$70 / mo
Nemotron 3.5 Lightning
$1.68 / mo
Performance comparison
Grok 4.7
46.4
Nemotron 3.5 Lightning
12.9
Overview
Pricing & capacity
| Pricing & capacity | Grok 4.7 | Nemotron 3.5 Lightning |
|---|---|---|
| Input / 1M tokens | $2.00 | $0.05 |
| Output / 1M tokens | $6.00 | $0.14 |
| Cached input / 1M | $0.50 | $0.02 |
| Monthly cost | $70 | $1.68 |
| Context window | 500K | 262K |
| Maximum output | 450K | 33K |
Input / 1M tokens
- Grok 4.7
- $2.00
- Nemotron 3.5 Lightning
- $0.05
Output / 1M tokens
- Grok 4.7
- $6.00
- Nemotron 3.5 Lightning
- $0.14
Cached input / 1M
- Grok 4.7
- $0.50
- Nemotron 3.5 Lightning
- $0.02
Monthly cost
- Grok 4.7
- $70
- Nemotron 3.5 Lightning
- $1.68
Context window
- Grok 4.7
- 500K
- Nemotron 3.5 Lightning
- 262K
Maximum output
- Grok 4.7
- 450K
- Nemotron 3.5 Lightning
- 33K
Benchmarks & performance
| Benchmarks & performance | Grok 4.7 | Nemotron 3.5 Lightning |
|---|---|---|
| Intelligence Index | 46.4 | 12.9 |
| GPQA | Not listed | 74.3% |
| Humanity’s Last Exam | 43.1% | 10.6% |
| SciCode | 57.4% | 32.1% |
| Output speed | 83 tokens/s | 301 tokens/s |
| Time to first token | 43.17 s | 0.56 s |
| DesignArena rating | 1,269 | Not listed |
| DesignArena win rate | 47.6% | Not listed |
Intelligence Index
- Grok 4.7
- 46.4
- Nemotron 3.5 Lightning
- 12.9
GPQA
- Grok 4.7
- Not listed
- Nemotron 3.5 Lightning
- 74.3%
Humanity’s Last Exam
- Grok 4.7
- 43.1%
- Nemotron 3.5 Lightning
- 10.6%
SciCode
- Grok 4.7
- 57.4%
- Nemotron 3.5 Lightning
- 32.1%
Output speed
- Grok 4.7
- 83 tokens/s
- Nemotron 3.5 Lightning
- 301 tokens/s
Time to first token
- Grok 4.7
- 43.17 s
- Nemotron 3.5 Lightning
- 0.56 s
DesignArena rating
- Grok 4.7
- 1,269
- Nemotron 3.5 Lightning
- Not listed
DesignArena win rate
- Grok 4.7
- 47.6%
- Nemotron 3.5 Lightning
- Not listed
Tools & features
| Tools & features | Grok 4.7 | Nemotron 3.5 Lightning |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Supported | Supported |
| Log probabilities | Supported | Supported |
| Deterministic seed | Supported | Supported |
| Prompt caching | Supported | Supported |
Function calling
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
Structured outputs
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
JSON mode
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
Reasoning
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
Log probabilities
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
Deterministic seed
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
Prompt caching
- Grok 4.7
- Supported
- Nemotron 3.5 Lightning
- Supported
API & availability
| API & availability | Grok 4.7 | Nemotron 3.5 Lightning |
|---|---|---|
| API identifier | grok-4.7 | nemotron-3.5-lightning |
| API providers | xAI, xAI, xAI, xAI, xAI | Not listed |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- Grok 4.7
grok-4.7- Nemotron 3.5 Lightning
nemotron-3.5-lightning
API providers
- Grok 4.7
- xAI, xAI, xAI, xAI, xAI
- Nemotron 3.5 Lightning
- Not listed
Prices checked
- Grok 4.7
- Oct 8, 2026
- Nemotron 3.5 Lightning
- Oct 8, 2026
About the models
Grok 4.7
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
Full pricing & detailsNemotron 3.5 Lightning
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Full pricing & detailsComparison FAQ
Grok 4.7: $2.00 input and $6.00 output per million tokens. Nemotron 3.5 Lightning: $0.05 input and $0.14 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
