Search results
Cutting-edge vision-language model exceling in high-quality reasoning from images.DownloadableFree Endpointllama-3.2-11b-vision-instruct
Cutting-edge vision-Language model exceling in high-quality reasoning from images.DownloadableFree Endpointllama-3.2-90b-vision-instruct
Leading content safety model for enhancing the safety and moderation capabilities of LLMsDownloadablellama-3.1-nemoguard-8b-content-safety
Topic control model to keep conversations focused on approved topics, avoiding inappropriate content.Downloadablellama-3.1-nemoguard-8b-topic-control
Leading multilingual content safety model for enhancing the safety and moderation capabilities of LLMsFree Endpointllama-3.1-nemotron-safety-guard-8b-v3
Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.DownloadableFree Endpointgemma-4-31b-it
Multimodal 320B-total / 18B-active MoE with hybrid KDA and sparse MLA attention, native FP8 weights, reasoning and tool calling.Free Endpointglm-5-3-flash
Open VLM for quantum computer calibration chart understanding across a range of qubit modalities.DownloadableFree Endpointising-calibration-1-35b-a3b
NVIDIA-Ising-Calibration-1.5 is a dense multimodal vision-language model built on Gemma 4 31B. It analyzes quantum computing calibration experiment plots and generates structured technical text.Free Endpointising-calibration-1.5-31b
Muse Glimmer 30B is a multimodal reasoning model accepting text and images, with native tool-calling and separate reasoning output.DownloadableFree Endpointmuse-glimmer-30b
1B embedding model for semantic search, retrieval, and RAG applications.Free Endpointnemotron-3-embed-1b
Nemotron 3 Nano Omni is an omni-modal reasoning model that understands images, video, speech, text.DownloadableFree Endpointnemotron-3-nano-omni-30b-a3b-reasoning
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and moreDownloadableFree Endpointnemotron-3-super-120b-a12b
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and moreDownloadableFree Endpointnemotron-3-ultra-550b-a55b
Multilingual, multimodal model for detecting unsafe and toxic content.DownloadableFree Endpointnemotron-3.5-content-safety
Fastest 30B A3B MoE model with leading domain accuracy for specialized agentic tasksDownloadableFree Endpointnemotron-3.5-lightning-30b-a3b
Stable Diffusion 3.5 is a popular text-to-image generation modelDownloadablestable-diffusion-3.5-large
Deploy and operate the RTVI-CV-3D microservice as MV3DT (`MODE=mv3dt`): per-camera DeepStream perception plus BEV Fusion over calibrated cameras. Supports the bundled sample dataset, custom video files, and RTSP streams, and chains to `vss-generate-video- Estimate 3D human body pose and skeleton from video input.DownloadableFree Endpoint3D Body Pose
Multi-modal model to classify safety for input prompts as well output responses.Free Endpointllama-guard-4-12b
One interface for supervised, RLHF, and parameter-efficient trainingPlaybooksAdvanced60 MINFine-Tune LLMs with LLaMA Factory
Multimodal question-answer retrieval representing user queries as text and documents as images.Downloadablellama-nemotron-embed-vl-1b-v2
GPU-accelerated model optimized for providing a probability score that a given passage contains the information to answer a question.Downloadablellama-nemotron-rerank-vl-1b-v2
Items per page
of 2 pages