Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

Models

Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

Optimized by NVIDIALaunch from Hugging FaceBeta

Filters (2)

  • Free Endpoint
    73
  • Partner Endpoint
    43
  • Download Available
    107
  • Drug Discovery
    13
  • Retrieval Augmented Generation
    10
  • Image-to-Text
    9
  • Speech-to-Text
    9
  • Image Generation
    8
  • Deepinfra
    33
  • OpenRouter
    28
  • Together AI
    21
  • GMI Cloud
    15
  • CoreWeave
    6
  • NVIDIA
    75
  • Meta
    11
  • Google
    6
  • Mistral AI
    6
  • Black forest labs
    4
  • H100 80GB HBM3
    12
  • A100 SXM4 80GB
    11
  • L40S
    11
  • A10G
    9
  • B200
    9
  • Chat
  • language generation
  • 1 model
    NVIDIA
    Free Endpoint

    nemotron-mini-4b-instruct

    Optimized SLM for on-device inference and fine-tuned for roleplay, RAG and function calling
    Chat
    2M
    1y
    Items per page
    of 1 pages