Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    Models

    Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

    Optimized by NVIDIALaunch from Hugging FaceBeta

    Filters (1)

    Use Case
    Inference Providers
    Publisher
    NIM Container GPUs
    Labels (1)
    4 models

    models

    • DeepSeek AI
      Free Endpoint

      deepseek-v4-flash-0731

      284B MoE (13B active) model ideal for long-context workloads optimized for coding, chat, and agentic workflows
      • MoE
      • Reasoning
      • Long Context
      • Hybrid Attention
      Last updated on August 19, 2026
    Items per page
    of 1 pages
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-ultra-550b-a55b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    • Agent
    • MoE
    • Frontier
    • Reasoning
    • Long Context
    52M API calls in the last 30 days
    Last updated on June 4, 2026
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-super-120b-a12b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    • MoE
    • Reasoning
    • Chat
    • Long Context
    • Instruction Following
    65M API calls in the last 30 days
    Last updated on March 11, 2026
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-30b-a3b

    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    • MoE
    • Reasoning
    • Long Context
    • Instruction Following
    12M API calls in the last 30 days
    Last updated on December 15, 2025