Skip to main content
Explore
Models
Skills
Blueprints
GPUs
Docs
Search
⌘K
Ctrl+K
?
Help Center
Getting Started
1
Set up your account
Create and verify your account to unlock full access to NVIDIA NIM APIs.
Create an Account
2
Generate API Key
3
Make your first API call
4
Prototype in your environment
5
Connect to inference partners
Resources
Developer Forums
Contact Support
FAQs
Login
8 results for
Filters (1)
Models (6)
Blueprints (1)
Skills (0)
Other (1)
Sort By
Best Match
Select item
Best Match
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Best Match
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Search results
NVIDIA
Downloadable
Video Super Resolution NIM
Upscale encoded or ST 2110 video to higher resolutions with NVIDIA Video Super Resolution.
Model
broadcast
+4
video upscaling
streaming
nvidia ai for media
video super resolution
1mo
Last updated on July 22, 2026
General
Launchable
Enterprise
Build a Video Search and Summarization (VSS) Agent
Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A
Blueprint
NVIDIA AI
+4
vision
video-to-text
generative AI
chat
7mo
Last updated on February 17, 2026
Items per page
24
12
24
48
96
1
1
of 1 pages
Wan-ai
Downloadable
wan2.2-animate-2-14b
Wan2.2-Animate-2 is a novel end-to-end character animation framework
Model
video editing
+1
character animation
27d
Last updated on August 19, 2026
NVIDIA
Downloadable
Free Endpoint
nemotron-3-nano-omni-30b-a3b-reasoning
Nemotron 3 Nano Omni is an omni-modal reasoning model that understands images, video, speech, text.
Model
Image-to-Text
+4
VLM
Video
Omni
OCR
8M
8M API calls in the last 30 days
4mo
Last updated on April 28, 2026
Playbooks
1 HR
Vision-Language Model Fine-tuning
Fine-tune Vision-Language Models for image and video understanding tasks using Qwen2.5-VL and InternVL3
Playbook
DGX
+6
Image Understanding
Vision-Language Models
GRPO
Spark
Fine-tuning
Video Analysis
11mo
Last updated on October 7, 2025
NVIDIA
Free Endpoint
cosmos-transfer2.5-2b
Generates physics-aware video world states for physical AI development using text prompts and multiple spatial control inputs derived from real-world data or simulation.
Model
Synthetic Data Generation
+4
Autonomous Vehicles
Physical AI
robotics
video-to-world
6mo
Last updated on February 26, 2026
Google
Free Endpoint
paligemma
Vision language model adept at comprehending text and visual inputs to produce informative responses
Model
image
+8
cv
Vision Assistant
vlm
Visual Question Answering
computer vision
Language Generation
video
Image-to-Text
12K
12K API calls in the last 30 days
2y
Last updated on August 26, 2024
NVIDIA
Downloadable
Free Endpoint
cosmos3-nano-reasoner
Vision language model that excels in understanding the physical world using structured reasoning on videos or images.
Model
video understanding
+8
autonomous vehicles
industrial
Physical AI
vision language model
reasoning
robotics
smart cities
Synthetic Data Generation
2K
2K API calls in the last 30 days
3mo
Last updated on June 1, 2026