NVIDIA
Explore
Models
Blueprints
GPUs
Docs
⌘KCtrl+K
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

Models

Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

Optimized by NVIDIALaunch from Hugging FaceBeta

Filters (2)

  • Download Available
    2
  • API Endpoint
    1
  • Code Generation
    0
  • Drug Discovery
    0
  • Image-to-Text
    0
  • Retrieval Augmented Generation
    0
  • Object Detection
    0
  • Qwen
    2
  • Z.ai
    1
  • NVIDIA
    0
  • Meta
    0
  • Mistral AI
    0
  • Agentic
  • MoE
  • 3 models
    Qwen
    qwen3.5-397b-a17b
    Next-gen Qwen 3.5 VLM (400B MoE) brings advanced vision, chat, RAG, and agentic capabilities.
    MoE
    4.66M
    2w
    Z.ai
    glm5
    GLM-5 744B MoE enables efficient reasoning for complex systems and long-horizon agentic tasks.
    MoE
    5.4M
    2w
    Qwen
    qwen3-coder-480b-a35b-instruct
    Excels in agentic coding and browser use and supports 256K context, delivering top results.
    agentic coding
    2.89M
    6mo
    Items per page
    of 1 pages