Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
ForumsSupport
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

Models

Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

Optimized by NVIDIALaunch from Hugging FaceBeta

Filters

Use Case
Inference Providers
Publisher
NIM Container GPUs
139 models
NVIDIA
Downloadable

Video Super Resolution NIM

Upscale encoded or ST 2110 video to higher resolutions with NVIDIA Video Super Resolution.
broadcast
video upscalingstreamingnvidia ai for mediavideo super resolution
Last updated on July 21, 2026
Items per page
of 6 pages
NVIDIA
Free Endpoint

nemotron-3-embed-1b

1B embedding model for semantic search, retrieval, and RAG applications.
Nemotron Retriever
Agentic RetrievalCode RetrievalText-to-EmbeddingRetrieval Augmented Generation
Last updated on July 16, 2026
Thinkingmachines
DownloadableFree Endpoint

inkling

Inkling is a multimodal (text + image) reasoning model from Thinking Machines — a Mamba-hybrid, 256-expert Mixture-of-Experts architecture with tool use and switchable reasoning.
text-to-text
reasoningimage-to-textmultimodal
Last updated on July 16, 2026
Poolside
Free Endpoint

laguna-xs-2.1

Efficient 33B MoE for local, long-horizon agentic coding and terminal tasks
Agentic AI
CodingReasoningTool Use
Last updated on July 15, 2026
Z.ai
DownloadableFree Endpoint

glm-5.2

GLM-5.2 is a flagship LLM for agentic workflows, coding, and long-horizon reasoning tasks.
Agentic AI
CodingReasoningTool Use
8M API calls in the last 30 days
Last updated on July 3, 2026
NVIDIA
Downloadable

qwen-image-edit-nvpcb-ovsl2sl

An image edit model specialized for Omniverse synthetic to photographic solder-light style captured at NVIDIA PCB inspection stations
Synthetic Data Generation
Image GenerationPhysical AI
Last updated on July 3, 2026
NVIDIA
Downloadable

nemotron-ocr-v2

Nemotron OCR v2 is a state-of-the-art multilingual text recognition model designed for robust end-to-end optical character recognition (OCR) on complex real-world images.
Table Extraction
nemo retrieverdata ingestionextractionOptical Character Recognition
338K API calls in the last 30 days
Last updated on June 24, 2026
Minimaxai
Free Endpoint

minimax-m3

MiniMax M3 Preview is a multimodal MoE vision-language model with strong reasoning, coding, and tool-calling capabilities.
coding
text-to-textreasoning
10M API calls in the last 30 days
Last updated on June 12, 2026
Google
DownloadableFree Endpoint

diffusiongemma-26b-a4b-it

Diffusion-based 26B parameter LLM enabling parallel token generation for real-time text apps
diffusion-llm
text-to-textreasoning
4M API calls in the last 30 days
Last updated on June 10, 2026
NVIDIA
DownloadableFree Endpoint

nemotron-3-ultra-550b-a55b

Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
Agent
MoEFrontierReasoningLong Context
52M API calls in the last 30 days
Last updated on June 4, 2026
Resemble.AI
Downloadable

chatterbox-multilingual-tts

Natural and expressive voices in 23 languages. For voice agents and brand ambassadors.
TTS
ChatterboxSpeech GenerationmultilingualText-to-Speech
7K API calls in the last 30 days
Last updated on June 3, 2026
NVIDIA
DownloadableFree Endpoint

nemotron-3.5-content-safety

Multilingual, multimodal model for detecting unsafe and toxic content.
llm safety
safety and moderationmultilingual content safetyai safety nemo guardrails
2M API calls in the last 30 days
Last updated on June 2, 2026
NVIDIA
Free Endpoint

cosmos3-nano

Generates physics-aware videos from text prompts or an image prompt for physical AI development.
autonomous vehicles
Physical AIroboticstext-to-worldimage-to-worldSynthetic Data Generation
2K API calls in the last 30 days
Last updated on June 1, 2026
NVIDIA
DownloadableFree Endpoint

cosmos3-nano-reasoner

Vision language model that excels in understanding the physical world using structured reasoning on videos or images.
video understanding
autonomous vehiclesindustrialPhysical AIvision language modelreasoningroboticssmart citiesSynthetic Data Generation
2K API calls in the last 30 days
Last updated on June 1, 2026
Stepfun-ai
DownloadableFree Endpoint

step-3.7-flash

A sparse MoE multimodal reasoning model good for enterprise, agentic and coding tasks.
Coding
VisionAgents
7M API calls in the last 30 days
Last updated on May 29, 2026
Moonshotai
DownloadableFree Endpoint

kimi-k2.6

1T multimodal MoE for long-horizon coding, agentic tool use, and image/video understanding.
Multimodal
Mixture-of-ExpertsReasoningImage-to-Text
16M API calls in the last 30 days
Last updated on May 1, 2026
Qwen
Downloadable

qwen-image

Qwen-Image is a text-to-image foundation model with advanced multilingual text rendering.
Text-to-Image
Image Generation
Last updated on May 1, 2026
Qwen
Downloadable

qwen-image-edit

Qwen-Image-Edit is an image editing model with multilingual text editing and strong subject consistency.
Text-to-Image
Image Generation
Last updated on May 1, 2026
Mistral AI
DownloadableFree Endpoint

mistral-medium-3.5-128b

A high performing model for text generation, coding and agentic use cases
coding
reasoningtextagentic
5M API calls in the last 30 days
Last updated on April 29, 2026
NVIDIA
DownloadableFree Endpoint

nemotron-3-nano-omni-30b-a3b-reasoning

Nemotron 3 Nano Omni is an omni-modal reasoning model that understands images, video, speech, text.
Image-to-Text
VLMVideoOmniOCR
8M API calls in the last 30 days
Last updated on April 28, 2026
DeepSeek AI
DownloadableFree Endpoint

deepseek-v4-flash

DeepSeek V4 Flash is a 284B MoE model with 1M-token context optimized for fast coding and agents.
MoE
codingfastagentic
15M API calls in the last 30 days
Last updated on April 24, 2026
DeepSeek AI
DownloadableFree Endpoint

deepseek-v4-pro

DeepSeek V4 scales to 1M-token context windows with efficient MoE architecture for coding tasks.
Moe
reasoningcodingagentic
8M API calls in the last 30 days
Last updated on April 24, 2026
NVIDIA
Deprecation in 7dDownloadable

Relighting

Re-illuminate people in video to match target lighting from a 360 HDRI environment map.
HDRI
remote contributionlightingnvidia ai for media
242 API calls in the last 30 days
Last updated on April 17, 2026
NVIDIA
DownloadableFree Endpoint

synthetic-video-detector

NVIDIA Synthetic Video Detector is an AI-powered micro-service for detecting AI‑generated (synthetic) videos.
broadcast
media2forensicsnvidia ai for mediadiffusion models
90K API calls in the last 30 days
Last updated on April 16, 2026