Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • DiscoverModelsSkillsBlueprintsGPUsDocsForums

    workstations

    • Run on RTX
    • Run on Spark
    • Run on Station

    models

    • Reasoning
    • Vision
    • Visual Design
    • Retrieval
    • Speech
    • Biology
    • Simulation
    • Climate & Weather
    • Safety & Moderation

    industries

    • Automotive
    • Financial Services
    • Gaming
    • Healthcare
    • Industrial
    • Robotics

    Vision

    Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    Deploy Models Now with NVIDIA NIM

    Optimized inference for the world’s leading models
    Free serverless APIs for development
    Accelerated by DGX Cloud
    Self-Host on your GPU infrastructure
    Continuous vulnerability fixes

    Explore NVIDIA Blueprints

    Comprehensive reference workflows that accelerate application development and deployment, featuring NVIDIA acceleration libraries, APIs, and microservices for AI agents, digital twins, and more.

    nvidiaBuild a Video Search and Summarization (VSS) Agent

    Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A

    • chat
    • generative AI
    • video-to-text
    • vision

    Vision Language Models (VLM)

    Multimodal models that can reason against image and video inputs and perform descriptive language generation​.

    Google
    DownloadableFree Endpoint

    diffusiongemma-26b-a4b-it

    Diffusion-based 26B parameter LLM enabling parallel token generation for real-time text apps
    • diffusion-llm
    • reasoning
    • text-to-text
    4M API calls in the last 30 days
    Last updated on June 10, 2026
    NVIDIA
    DownloadableFree Endpoint

    cosmos3-nano-reasoner

    Vision language model that excels in understanding the physical world using structured reasoning on videos or images.
    • Physical AI
    • autonomous vehicles
    • industrial
    • reasoning
    • robotics
    • smart cities
    • video understanding
    • vision language model
    2K API calls in the last 30 days
    Last updated on June 1, 2026
    Google
    Free Endpoint

    paligemma

    Vision language model adept at comprehending text and visual inputs to produce informative responses
    • Language Generation
    • Vision Assistant
    • Visual Question Answering
    • computer vision
    • cv
    • image
    • video
    • vlm
    12K API calls in the last 30 days
    Last updated on August 26, 2024