Nvidia

Nvidia

Nvidia is an AI model provider.Tokenando tracks 17 Nvidia models, with input pricing from $0.040/M and an average blended cost of $0.831/M. Its flagship model is Llama-3.1-Nemotron-Ultra-253B.

NIM endpoints expose Nemotron and partner models behind a uniform API. Tight integration with NVIDIA enterprise stack.

Founded 1993HQ Santa Clara, USAWebsite ↗API docs ↗
MODELS TRACKED
17
3 categories
FLAGSHIP
Llama-3.1-Nemotron-Ultra-253B
Live API
MIN INPUT
$0.040/M
cheapest model in family
AVG BLENDED
$0.831/M
across 16 priced models
MAX CONTEXT
1,000,000
largest window in family

Frontier

3 models

Multimodal

3 models

Efficient

11 models

Frequently Asked Questions

How many models does Nvidia offer?

Tokenando tracks 17 Nvidia models.

How much do Nvidia models cost?

Nvidia model input pricing starts at $0.040 per million tokens, with an average blended cost of $0.831 per million across the 16 priced models we track.

What is Nvidia's flagship model?

Nvidia's flagship model is Llama-3.1-Nemotron-Ultra-253B. It is the highest-tier Nvidia model we track, with input pricing of $1.600 per million tokens.

What model categories does Nvidia cover?

Nvidia covers 3 categories: frontier, multimodal and efficient.

Terms used on this page