Models37in catalog
Largest context1MNemotron 3.5 Lightning (free)
Avg context
278K
Cheapest input
Free/M
Nemotron 3.5 Lightning (free)
Largest output
262K
Fw Nemotron 3 Ultra Nvfp4
Speed tiers
balanced 18fast 18deep 1
Capabilities
What Nvidia models support across the catalog.
Vision
10 / 37
27% of models
Tool use
28 / 37
76% of models
Function calling
28 / 37
76% of models
Extended thinking
28 / 37
76% of models
Streaming
37 / 37
100% of models
Prompt caching
11 / 37
30% of models
All Nvidia models
Sorted by context window, largest first. Tap a row for full specs.
Showing 25 of 37 models
See rankings
How Nvidia models rank against the full catalog.