Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

12 results for

Filters (1)

  • Free Endpoint
    12
  • Partner Endpoint
    5
  • Download Available
    10
  • Developer Example
    0
  • Launchable
    0
  • Code Generation
    1
  • Image-to-Text
    0
  • Retrieval Augmented Generation
    0
  • Text-to-Embedding
    0
  • Image Generation
    0
  • Deepinfra
    4
  • Together AI
    3
  • GMI Cloud
    3
  • Digital Ocean
    3
  • Bitdeer
    2
  • NVIDIA
    7
  • Meta
    1
  • Google
    1
  • Mistral AI
    1
  • Stepfun ai
    1
  • Application Developer
    0
  • DevOps Engineer
    0
  • Platform Engineer
    0
  • Solutions Architect
    0
  • NVIDIA AI
    0
  • AI And Machine Learning
    0
  • H200
    5
  • L40S
    5
  • B200
    4
  • H100 80GB HBM3
    4
  • A100 SXM4 80GB
    4
  • Video Search and Summarization (VSS)
    0
  • reasoning
  • Meta
    DownloadableFree Endpoint

    llama-3.3-70b-instruct

    Advanced LLM for reasoning, math, general knowledge, and function calling
    Model
    Instruction following
    Items per page
    of 1 pages
    18.79M
    1y
    NVIDIA
    DownloadableFree Endpoint

    llama-3.1-nemotron-nano-8b-v1

    Leading reasoning and agentic AI accuracy model for PC and edge.
    Model
    advanced reasoning
    1.47M
    11mo
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    4.93M
    11mo
    NVIDIA
    DownloadableFree Endpoint

    llama-3.3-nemotron-super-49b-v1.5

    High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
    Model
    advanced reasoning
    3.17M
    10mo
    Google
    DownloadableFree Endpoint

    gemma-4-31b-it

    Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.
    Model
    reasoning
    5.49M
    2mo
    NVIDIA
    DownloadableFree Endpoint

    ising-calibration-1-35b-a3b

    Open VLM for quantum computer calibration chart understanding across a range of qubit modalities.
    Model
    Quantum
    332K
    2mo
    Mistral AI
    DownloadableFree Endpoint

    mistral-medium-3.5-128b

    A high performing model for text generation, coding and agentic use cases
    Model
    coding
    3.76M
    1mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-30b-a3b

    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    Model
    MoE
    11.91M
    6mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-super-120b-a12b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    MoE
    60.41M
    3mo
    NVIDIA
    DownloadableFree Endpoint

    nemotron-3-ultra-550b-a55b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    Agent
    7.73M
    11d
    ByteDance
    Free Endpoint

    seed-oss-36b-instruct

    ByteDance open-source LLM with long-context, reasoning, and agentic intelligence.
    Model
    thinking budget
    1.18M
    9mo
    Stepfun-ai
    Free Endpoint

    step-3.5-flash

    200B open-source reasoning engine with sparse MoE powering frontier agentic AI.
    Model
    Agentic
    11.71M
    4mo