Model provider
NVIDIA
NIM-hosted Nemotron and open models optimized for accelerated inference.
chat
Available through one endpoint
API documentation →5 NVIDIA models
NVIDIAnorth_eastNemotron 3 Ultra
nemotron-ultraNVIDIA flagship 550B MoE reasoning model served through NIM.
reasoningopen-weights NVIDIAnorth_eastNemotron 3 Supernemotron-3-superMid-size NVIDIA Nemotron MoE for reasoning and agents on NIM.
reasoningtoolsopen-weights NVIDIAnorth_eastNemotron 3 Nanonemotron-3-nanoLow-cost NVIDIA Nemotron MoE for fast chat and extraction on NIM.
fastlow-costopen-weights NVIDIAnorth_eastLlama Nemotron Super 49Bnemotron-super-49bLlama-based Nemotron tuned for reasoning, served through NIM.
reasoningopen-weights NVIDIAnorth_eastGPT-OSS 120B on NVIDIAnvidia-gpt-oss-120bOpenAI open-weights 120B served through NIM.
chatcodeopen-weights