Skip to main content
Explore
Models
Skills
Blueprints
GPUs
Docs
Search
⌘K
Ctrl+K
?
Help Center
Getting Started
1
Set up your account
Create and verify your account to unlock full access to NVIDIA NIM APIs.
Create an Account
2
Generate API Key
3
Make your first API call
4
Prototype in your environment
5
Connect to inference partners
Resources
Developer Forums
Contact Support
FAQs
Login
Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
Optimized by NVIDIA
Launch from Hugging Face
Beta
Filters
11 models
Sort By
Most Recent
Select item
Most Recent
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Most Recent
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
models
NVIDIA
Downloadable
Free Endpoint
nemotron-3.5-lightning-30b-a3b
Fastest 30B A3B MoE model with leading domain accuracy for specialized agentic tasks
Customization
+3
Text-to-Text
Long-running agents
Open
25d
Last updated on August 11, 2026
Items per page
24
12
24
48
96
1
1
of 1 pages
Meta
Downloadable
Free Endpoint
muse-glimmer-30b
Muse Glimmer 30B is a multimodal reasoning model accepting text and images, with native tool-calling and separate reasoning output.
Multimodal
+5
Image-to-Text
Reasoning
Chat
Text-to-Text
Large Language Models
26d
Last updated on August 10, 2026
Minimaxai
Deprecation in 4d
Free Endpoint
minimax-m3
MiniMax M3 Preview is a multimodal MoE vision-language model with strong reasoning, coding, and tool-calling capabilities.
coding
+2
text-to-text
reasoning
10M
10M API calls in the last 30 days
2mo
Last updated on June 12, 2026
Google
Downloadable
Free Endpoint
diffusiongemma-26b-a4b-it
Diffusion-based 26B parameter LLM enabling parallel token generation for real-time text apps
diffusion-llm
+2
text-to-text
reasoning
4M
4M API calls in the last 30 days
2mo
Last updated on June 10, 2026
NVIDIA
Downloadable
Free Endpoint
cosmos3-nano
Generates physics-aware videos from text prompts or an image prompt for physical AI development.
autonomous vehicles
+5
Physical AI
robotics
text-to-world
image-to-world
Synthetic Data Generation
2K
2K API calls in the last 30 days
3mo
Last updated on June 1, 2026
Qwen
Downloadable
qwen-image
Qwen-Image is a text-to-image foundation model with advanced multilingual text rendering.
Text-to-Image
+1
Image Generation
4mo
Last updated on May 1, 2026
Google
Downloadable
Free Endpoint
gemma-4-31b-it
Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.
reasoning
+3
coding
text-to-text
agentic
6M
6M API calls in the last 30 days
5mo
Last updated on April 2, 2026
Microsoft
Downloadable
TRELLIS
MSFT TRELLIS is a 3D AI model that generates high-quality 3D assets from text or image inputs.
text-to-3d
+2
Run-on-RTX
image-to-3d
18K
18K API calls in the last 30 days
1y
Last updated on September 3, 2025
Stability AI
Downloadable
stable-diffusion-3.5-large
Stable Diffusion 3.5 is a popular text-to-image generation model
Text-to-Image
+1
Image Generation
1y
Last updated on August 12, 2025
OpenAI
Downloadable
Free Endpoint
gpt-oss-20b
Smaller Mixture of Experts (MoE) text-only LLM for efficient AI reasoning and math
reasoning
+3
text-to-text
chat
math
19M
19M API calls in the last 30 days
1y
Last updated on August 5, 2025
NVIDIA
Free Endpoint
magpie-tts-zeroshot
Expressive and engaging text-to-speech, generated from a short audio sample.
TTS
+3
NVIDIA NIM
NVIDIA Riva
Text-to-Speech
17K
17K API calls in the last 30 days
1y
Last updated on June 12, 2025