Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

29 results for

Filters (1)

  • Free Endpoint
    28
  • Partner Endpoint
    20
  • Download Available
    25
  • Launchable
    0
  • Developer Example
    0
  • Enterprise Blueprint
    0
  • NemoClaw Blueprint
    0
  • Code Generation
    2
  • Synthetic Data Generation
    2
  • Image-to-Text
    1
  • Drug Discovery
    0
  • Retrieval Augmented Generation
    0
  • OpenRouter
    16
  • Deepinfra
    15
  • Together AI
    12
  • GMI Cloud
    10
  • Bitdeer
    5
  • NVIDIA
    10
  • Mistral AI
    3
  • Google
    2
  • OpenAI
    2
  • Minimaxai
    2
  • Developer
    0
  • AI Engineer
    0
  • Ml Engineer
    0
  • Application Developer
    0
  • Platform Engineer
    0
  • NVIDIA AI
    0
  • NVIDIA Omniverse
    0
  • NVIDIA BioNemo
    0
  • NVIDIA Isaac GR00T
    0
  • AI And Machine Learning
    0
  • Physical AI
    0
  • Accelerated Computing
    0
  • Developer Tools
    0
  • Infrastructure
    0
  • H100 80GB HBM3
    4
  • A100 SXM4 80GB
    4
  • L40S
    4
  • A10G
    4
  • H200
    4
  • TAO Toolkit
    0
  • Jetson
    0
  • NeMo Megatron Bridge
    0
  • Video Search and Summarization (VSS)
    0
  • MONAI
    0
  • reasoning
  • NVIDIA
    Downloadable

    cosmos-reason2-8b

    Vision language model that excels in understanding the physical world using structured reasoning on videos or images.
    Model
    video understanding
    Items per page
    of 2 pages
    191K
    6mo
    NVIDIA
    DownloadableFree Endpoint

    cosmos3-nano-reasoner

    Vision language model that excels in understanding the physical world using structured reasoning on videos or images.
    Model
    video understanding
    2K
    1mo
    DeepSeek AI
    DownloadableFree Endpoint

    deepseek-v4-pro

    DeepSeek V4 scales to 1M-token context windows with efficient MoE architecture for coding tasks.
    Model
    Moe
    8M
    2mo
    Google
    DownloadableFree Endpoint

    diffusiongemma-26b-a4b-it

    Diffusion-based 26B parameter LLM enabling parallel token generation for real-time text apps
    Model
    diffusion-llm
    4M
    1mo
    Google
    DownloadableFree Endpoint

    gemma-4-31b-it

    Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.
    Model
    reasoning
    6M
    3mo
    Z.ai
    DownloadableFree Endpoint

    glm-5.2

    GLM-5.2 is a flagship LLM for agentic workflows, coding, and long-horizon reasoning tasks.
    Model
    Agentic AI
    8M
    16d
    OpenAI
    DownloadableFree Endpoint

    gpt-oss-120b

    Mixture of Experts (MoE) reasoning LLM (text-only) designed to fit within 80GB GPU.
    Model
    reasoning
    45M
    11mo
    OpenAI
    DownloadableFree Endpoint

    gpt-oss-20b

    Smaller Mixture of Experts (MoE) text-only LLM for efficient AI reasoning and math
    Model
    reasoning
    18M
    11mo
    Thinkingmachines
    DownloadableFree Endpoint

    inkling

    Inkling is a multimodal (text + image) reasoning model from Thinking Machines — a Mamba-hybrid, 256-expert Mixture-of-Experts architecture with tool use and switchable reasoning.
    Model
    text-to-text
    3d
    NVIDIA
    DownloadableFree Endpoint

    ising-calibration-1-35b-a3b

    Open VLM for quantum computer calibration chart understanding across a range of qubit modalities.
    Model
    Quantum
    442K
    3mo
    Moonshotai
    DownloadableFree Endpoint

    kimi-k2.6

    1T multimodal MoE for long-horizon coding, agentic tool use, and image/video understanding.
    Model
    Multimodal
    16M
    2mo
    Poolside
    Free Endpoint

    laguna-xs-2.1

    Efficient 33B MoE for local, long-horizon agentic coding and terminal tasks
    Model
    Agentic AI
    3d
    NVIDIA
    DownloadableFree Endpoint

    llama-3.1-nemotron-nano-8b-v1

    Leading reasoning and agentic AI accuracy model for PC and edge.
    Model
    advanced reasoning
    1M
    1y
    Meta
    DownloadableFree Endpoint

    llama-3.3-70b-instruct

    Advanced LLM for reasoning, math, general knowledge, and function calling
    Model
    Instruction following
    19M
    1y
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    5M
    1y
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1.5

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    3M
    11mo
    Minimaxai
    DownloadableFree Endpoint

    minimax-m2.7

    MiniMax M2.7 is a 230B-parameter text-to-text AI model excelling in coding, reasoning, and office tasks.
    Model
    coding
    16M
    3mo
    Minimaxai
    Free Endpoint

    minimax-m3

    MiniMax M3 Preview is a multimodal MoE vision-language model with strong reasoning, coding, and tool-calling capabilities.
    Model
    coding
    10M
    1mo
    Mistral AI
    DownloadableFree Endpoint

    mistral-medium-3.5-128b

    A high performing model for text generation, coding and agentic use cases
    Model
    coding
    5M
    2mo
    Mistral AI
    DownloadableFree Endpoint

    mistral-small-4-119b-2603

    Hybrid MoE model unifying instruct, reasoning, and coding with multimodal input and 256k context
    Model
    code generation
    11M
    4mo
    Mistral AI
    DownloadableFree Endpoint

    mixtral-8x7b-instruct-v0.1

    An MOE LLM that follows instructions, completes requests, and generates creative text.
    Model
    Advanced Reasoning
    1M
    1y
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-30b-a3b

    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    Model
    MoE
    12M
    7mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-super-120b-a12b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    MoE
    60M
    4mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-ultra-550b-a55b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    Agent
    52M
    1mo