Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

27 results for

Filters (1)

  • Free Endpoint
    27
  • Partner Endpoint
    16
  • Download Available
    20
  • Developer Example
    0
  • Launchable
    0
  • Code Generation
    5
  • Image-to-Text
    5
  • Retrieval Augmented Generation
    0
  • Text-to-Embedding
    0
  • Image Generation
    0
  • Deepinfra
    10
  • Together AI
    8
  • Bitdeer
    7
  • GMI Cloud
    6
  • CoreWeave
    4
  • NVIDIA
    9
  • Meta
    8
  • Google
    3
  • Mistral AI
    2
  • Stepfun ai
    2
  • Application Developer
    0
  • DevOps Engineer
    0
  • Platform Engineer
    0
  • Solutions Architect
    0
  • NVIDIA AI
    0
  • AI And Machine Learning
    0
  • H200
    10
  • B200
    9
  • H100 80GB HBM3
    9
  • L40S
    9
  • A100 SXM4 80GB
    8
  • Video Search and Summarization (VSS)
    0
  • Chat
  • Meta
    DownloadableFree Endpoint

    llama-3.1-70b-instruct

    Powers complex conversations with superior contextual understanding, reasoning and text generation.
    Model
    Chat
    Items per page
    of 2 pages
    3.9M
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.1-8b-instruct

    Advanced state-of-the-art model with language understanding, superior reasoning, and text generation.
    Model
    Chat
    25.09M
    11mo
    Meta
    DownloadableFree Endpoint

    llama-3.2-1b-instruct

    Advanced state-of-the-art small language model with language understanding, superior reasoning, and text generation.
    Model
    Language Generation
    46.31K290K
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.2-3b-instruct

    Advanced state-of-the-art small language model with language understanding, superior reasoning, and text generation.
    Model
    Language Generation
    27.53K1.22M
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.3-70b-instruct

    Advanced LLM for reasoning, math, general knowledge, and function calling
    Model
    Instruction following
    18.79M
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.2-11b-vision-instruct

    Cutting-edge vision-language model exceling in high-quality reasoning from images.
    Model
    Image-Text Retrieval
    1.67M
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.2-90b-vision-instruct

    Cutting-edge vision-Language model exceling in high-quality reasoning from images.
    Model
    Image-Text Retrieval
    2.69M
    1y
    NVIDIA
    DownloadableFree Endpoint

    llama-3.1-nemotron-nano-8b-v1

    Leading reasoning and agentic AI accuracy model for PC and edge.
    Model
    advanced reasoning
    1.47M
    11mo
    NVIDIA
    DownloadableFree Endpoint

    llama-3.1-nemotron-nano-vl-8b-v1

    Multi-modal vision-language model that understands text/img and creates informative responses
    Model
    doc intelligence
    10.15M
    11mo
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    4.93M
    11mo
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1.5

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    3.17M
    10mo
    Abacus.AI
    Free Endpoint

    dracarys-llama-3.1-70b-instruct

    Fine-tuned Llama 3.1 70B model for code generation, summarization, and multi-language tasks.
    Model
    Code Generation
    783K
    1y
    Google
    Free Endpoint

    gemma-3n-e2b-it

    An edge computing AI model which accepts text, audio and image input, ideal for resource-constrained environments
    Model
    language generation
    33.75M
    11mo
    Google
    Free Endpoint

    gemma-3n-e4b-it

    An edge computing AI model which accepts text, audio and image input, ideal for resource-constrained environments
    Model
    language generation
    3.79M
    11mo
    Google
    DownloadableFree Endpoint

    gemma-4-31b-it

    Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.
    Model
    reasoning
    5.49M
    2mo
    NVIDIA
    DownloadableFree Endpoint

    ising-calibration-1-35b-a3b

    Open VLM for quantum computer calibration chart understanding across a range of qubit modalities.
    Model
    Quantum
    332K
    2mo
    Mistral AI
    Free Endpoint

    mistral-large-3-675b-instruct-2512

    A state-of-the-art general purpose MoE VLM ideal for chat, agentic and instruction based use cases.
    Model
    language generation
    3.22M
    6mo
    Mistral AI
    DownloadableFree Endpoint

    mistral-medium-3.5-128b

    A high performing model for text generation, coding and agentic use cases
    Model
    coding
    3.76M
    1mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-30b-a3b

    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    Model
    MoE
    11.91M
    6mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-omni-30b-a3b-reasoning

    Nemotron 3 Nano Omni is an omni-modal reasoning model that understands images, video, speech, text.
    Model
    Image-to-Text
    7.54M
    1mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-super-120b-a12b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    MoE
    60.41M
    3mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-ultra-550b-a55b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    Agent
    7.73M
    11d
    Qwen
    DownloadableFree Endpoint

    qwen3.5-397b-a17b

    Next-gen Qwen 3.5 VLM (400B MoE) brings advanced vision, chat, RAG, and agentic capabilities.
    Model
    MoE
    13.15M
    3mo
    ByteDance
    Free Endpoint

    seed-oss-36b-instruct

    ByteDance open-source LLM with long-context, reasoning, and agentic intelligence.
    Model
    thinking budget
    1.18M
    9mo