NVIDIA
Explore
Models
Blueprints
GPUs
Docs
⌘KCtrl+K
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

Models

Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

Optimized by NVIDIALaunch from Hugging FaceBeta

Filters (2)

  • API Endpoint
    4
  • Download Available
    1
  • Code Generation
    0
  • Drug Discovery
    0
  • Image-to-Text
    0
  • Retrieval Augmented Generation
    0
  • Object Detection
    0
  • Moonshotai
    2
  • NVIDIA
    1
  • Qwen
    1
  • DeepSeek AI
    1
  • Meta
    0
  • long context
  • long-context
  • 5 models
    DeepSeek AI
    deepseek-v3.2
    State-of-the-art 685B reasoning LLM with sparse attention, long context, and integrated agentic tools.
    long context
    13.91M
    2mo
    NVIDIA
    nemotron-3-nano-30b-a3b
    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    MoE
    10.59M
    2mo
    Moonshotai
    kimi-k2-thinking
    Open reasoning model with 256K context window, native INT4 quantization and enhanced tool use.
    Conversational
    2.83M
    2mo
    Moonshotai
    kimi-k2-instruct-0905
    Follow-on version of Kimi-K2-Instruct with longer context window and enhanced reasoning capabilities.
    long-context
    10.27M
    5mo
    Qwen
    qwen3-coder-480b-a35b-instruct
    Excels in agentic coding and browser use and supports 256K context, delivering top results.
    agentic coding
    2.89M
    6mo
    Items per page
    of 1 pages