Nemotron 3 Ultra vs Mistral Large 3 2512 compared — price per token, context window, multimodality, openness and which to choose. Full 2026 breakdown.
Prices & specs refreshed from live data · olud.ai
Open-model prices = cheapest provider via OpenRouter; official maker rates may be higher.
| Spec | Nemotron 3 Ultra | Mistral Large 3 2512 |
|---|---|---|
| Maker | NVIDIA | Mistral AI |
| Type | Open-weight | Open-weight |
| Context window | 512K tokens | 262K tokens |
| Input price | $0.6/M · free self-host | $0.5/M · free self-host |
| Output price | $3.6/M · free self-host | $1.5/M · free self-host |
| Vision / multimodal | No | Yes |
| Tool / function calling | Yes | Yes |
| Self-hostable | Yes | Yes |
| License | NVIDIA Open Model | Apache 2.0 |
| Capability | Nemotron 3 Ultra | Mistral Large 3 2512 |
|---|---|---|
| Open weights (downloadable) | ✓ | ✓ |
| Self-hostable | ✓ | ✓ |
| Runs fully offline | ✓ | ✓ |
| Vision / multimodal | ✗ | ✓ |
| Tool / function calling | ✓ | ✓ |
| 1M+ context window | ✗ | ✗ |
Independent benchmark scores measured by Artificial Analysis. Higher is better (except latency).
Benchmark data by Artificial Analysis.
| Criterion | Nemotron 3 Ultra | Mistral Large 3 2512 |
|---|---|---|
| Cost-efficiency | 4.0 | 4.5 |
| Context window | 4.5 | 4.0 |
| Openness | 5.0 | 5.0 |
| Self-hosting | 5.0 | 5.0 |
| Multimodality | 3.5 | 5.0 |
Scores come from live data — output price (cost), context length, open vs closed weights (openness & self-hosting) and vision/tool support (multimodality). Raw task quality isn't scored here; it depends on your benchmark — see the verdict.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
These variants are tracked but not compared here — one page per family keeps the comparison readable.
Mistral Large 3 2512 is cheaper on output ($1.5/M vs $3.6/M).
Nemotron 3 Ultra offers the larger context window (512K tokens).
Choose Mistral Large 3 2512 for the lower output price ($1.5/M vs $3.6/M). Choose Nemotron 3 Ultra if you need the larger 512K context window.
Browse the full open-source model leaderboard, thousands of tools and live benchmarks — all in one place.
Open the leaderboard →