Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

Models

Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices

Optimized by NVIDIALaunch from Hugging FaceBeta

Filters (1)

  • Free Endpoint
    3
  • Partner Endpoint
    0
  • Download Available
    2
  • Image-to-Text
    3
  • Drug Discovery
    0
  • Retrieval Augmented Generation
    0
  • Speech-to-Text
    0
  • Image Generation
    0
  • Deepinfra
    0
  • OpenRouter
    0
  • Together AI
    0
  • GMI Cloud
    0
  • Bitdeer
    0
  • Mistral AI
    2
  • NVIDIA
    1
  • Google
    1
  • Meta
    0
  • Qwen
    0
  • H100 80GB HBM3
    0
  • A100 SXM4 80GB
    0
  • L40S
    0
  • A10G
    0
  • B200
    0
  • language generation
  • 4 models
    Mistral AI
    Downloadable

    mistral-large-3-675b-instruct-2512

    A state-of-the-art general purpose MoE VLM ideal for chat, agentic and instruction based use cases.
    language generation
    2M
    7mo
    Items per page
    of 1 pages
    Mistral AI
    DownloadableFree Endpoint

    ministral-14b-instruct-2512

    A general purpose VLM ideal for chat and instruction based use cases
    language generation
    4M
    7mo
    NVIDIA
    Free Endpoint

    nemotron-mini-4b-instruct

    Optimized SLM for on-device inference and fine-tuned for roleplay, RAG and function calling
    Chat
    2M
    1y
    Google
    Free Endpoint

    paligemma

    Vision language model adept at comprehending text and visual inputs to produce informative responses
    image
    12K
    1y