Explore
Models
Blueprints
GPUs
Docs
⌘K
Ctrl+K
?
Login
30 results for
Filters (1)
Models (29)
Blueprints (1)
Other (0)
Sort By
score:DESC
Best Match
NVIDIA
Free Endpoint
nemotron-voicechat
Nemotron 3 Voicechat
Model
English
+2
1mo
Items per page
24
1
1
2
2
of 2 pages
2.65K
Google
Free Endpoint
gemma-2-2b-it
Advanced small language generative AI model for edge applications
Model
Chat
+3
571K
11mo
Google
Free Endpoint
gemma-3n-e2b-it
An edge computing AI model which accepts text, audio and image input, ideal for resource-constrained environments
Model
language generation
+3
5.02M
9mo
Moonshotai
Deprecation in 3d
Free Endpoint
kimi-k2-instruct
State-of-the-art open mixture-of-experts model with strong reasoning, coding, and agentic capabilities
Model
coding
+3
10.35M
9mo
NVIDIA
Free Endpoint
nemotron-mini-4b-instruct
Optimized SLM for on-device inference and fine-tuned for roleplay, RAG and function calling
Model
Chat
+2
912K
1y
NVIDIA
Free Endpoint
usdcode
State-of-the-art LLM that answers OpenUSD knowledge queries and generates USD-Python code.
Model
Digital Twin
+4
10mo
NVIDIA
Launchable
Enterprise
Build a Video Search and Summarization (VSS) Agent
Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A
Blueprint
NVIDIA AI
+4
2mo
Google
Free Endpoint
gemma-3n-e4b-it
An edge computing AI model which accepts text, audio and image input, ideal for resource-constrained environments
Model
language generation
+3
2.24M
9mo
OpenAI
Downloadable
gpt-oss-120b
Mixture of Experts (MoE) reasoning LLM (text-only) designed to fit within 80GB GPU.
Model
reasoning
+3
32.44M
9mo
Meta
Downloadable
llama-3.1-70b-instruct
Powers complex conversations with superior contextual understanding, reasoning and text generation.
Model
Chat
+3
2.98M
11mo
Meta
Downloadable
llama-3.2-1b-instruct
Advanced state-of-the-art small language model with language understanding, superior reasoning, and text generation.
Model
chat
+3
28.38K
422K
11mo
Meta
Downloadable
llama-3.2-3b-instruct
Advanced state-of-the-art small language model with language understanding, superior reasoning, and text generation.
Model
Chat
+3
18.42K
1.13M
11mo
Mistral AI
Downloadable
mistral-7b-instruct-v0.3
This LLM follows instructions, completes requests, and generates creative text.
Model
Chat
+2
416K
11mo
Mistral AI
Downloadable
mixtral-8x22b-instruct-v0.1
An MOE LLM that follows instructions, completes requests, and generates creative text.
Model
Advanced Reasoning
+4
2.21M
9mo
Mistral AI
Downloadable
mixtral-8x7b-instruct-v0.1
An MOE LLM that follows instructions, completes requests, and generates creative text.
Model
Advanced Reasoning
+4
648K
9mo
NVIDIA
Downloadable
nemotron-3-super-120b-a12b
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
Model
MoE
+4
50.65M
2mo
Microsoft
Downloadable
phi-4-mini-instruct
Lightweight multilingual LLM powering AI applications in latency bound, memory/compute constrained environments
Model
Chat
+3
606K
11mo
Upstage
Free Endpoint
solar-10.7b-instruct
Excels in NLP tasks, particularly in instruction-following, reasoning, and mathematics.
Model
Non-Commercial Use Only
+4
325K
1y
OpenAI
Downloadable
gpt-oss-20b
Smaller Mixture of Experts (MoE) text-only LLM for efficient AI reasoning and math
Model
reasoning
+3
14.08M
9mo
Meta
Downloadable
llama-3.1-8b-instruct
Advanced state-of-the-art model with language understanding, superior reasoning, and text generation.
Model
Chat
+4
23.03M
10mo
NVIDIA
Downloadable
llama-3.3-nemotron-super-49b-v1
High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
Model
math
+3
3.42M
9mo
NVIDIA
Downloadable
llama-3.3-nemotron-super-49b-v1.5
High efficiency model with leading accuracy for reasoning, tool calling, chat, and instruction following.
Model
math
+3
3.06M
9mo
Mistral AI
Downloadable
ministral-14b-instruct-2512
A general purpose VLM ideal for chat and instruction based use cases
Model
language generation
+3
2.45M
5mo
Qwen
Downloadable
qwen3.5-122b-a10b
122B MoE LLM (10B active) for coding, reasoning, multimodal chat. Agent-ready.
Model
tool calling
+3
10.72M
2mo