Skip to main content
Explore
Models
Skills
Blueprints
GPUs
Docs
Search
⌘K
Ctrl+K
?
Forums
Support
Login
Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
Optimized by NVIDIA
Launch from Hugging Face
Beta
Filters (1)
5 models
Sort By
Most Recent
Select item
Most Recent
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Most Recent
Most Popular
Most Downloaded
Alphabetical (A-Z)
Alphabetical (Z-A)
Thinkingmachines
Downloadable
Free Endpoint
inkling
Inkling is a multimodal (text + image) reasoning model from Thinking Machines — a Mamba-hybrid, 256-expert Mixture-of-Experts architecture with tool use and switchable reasoning.
text-to-text
+3
reasoning
image-to-text
multimodal
5d
Last updated on July 16, 2026
Items per page
24
12
24
48
96
1
1
of 1 pages
Moonshotai
Downloadable
Free Endpoint
kimi-k2.6
1T multimodal MoE for long-horizon coding, agentic tool use, and image/video understanding.
Multimodal
+3
Mixture-of-Experts
Reasoning
Image-to-Text
16M
16M API calls in the last 30 days
2mo
Last updated on May 1, 2026
Mistral AI
Downloadable
mistral-large-3-675b-instruct-2512
A state-of-the-art general purpose MoE VLM ideal for chat, agentic and instruction based use cases.
language generation
+3
multimodal
agentic
Image-to-Text
2M
2M API calls in the last 30 days
7mo
Last updated on December 2, 2025
Mistral AI
Downloadable
Free Endpoint
ministral-14b-instruct-2512
A general purpose VLM ideal for chat and instruction based use cases
language generation
+3
SLM
multimodal
Image-to-Text
4M
4M API calls in the last 30 days
7mo
Last updated on December 2, 2025
Meta
Free Endpoint
llama-guard-4-12b
Multi-modal model to classify safety for input prompts as well output responses.
LLM Multimodal Safety
+3
Content Safety
Guardrail
Content Moderator
357K
357K API calls in the last 30 days
1y
Last updated on July 1, 2025