Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
models
Cutting-edge vision-language model excelling in retrieving text and metadata from images.Downloadablenemotron-parse-2.0
Muse Glimmer 30B is a multimodal reasoning model accepting text and images, with native tool-calling and separate reasoning output.DownloadableFree Endpointmuse-glimmer-30b
GPU-accelerated model optimized for providing a probability score that a given passage contains the information to answer a question.Downloadablellama-nemotron-rerank-vl-1b-v2
Multimodal question-answer retrieval representing user queries as text and documents as images.Downloadablellama-nemotron-embed-vl-1b-v2
Cutting-edge vision-language model exceling in retrieving text and metadata from images.Downloadablenemotron-parse
Leading multilingual content safety model for enhancing the safety and moderation capabilities of LLMsFree Endpointllama-3.1-nemotron-safety-guard-8b-v3
Multi-modal model to classify safety for input prompts as well output responses.Free Endpointllama-guard-4-12b
Topic control model to keep conversations focused on approved topics, avoiding inappropriate content.Downloadablellama-3.1-nemoguard-8b-topic-control
Leading content safety model for enhancing the safety and moderation capabilities of LLMsDownloadablellama-3.1-nemoguard-8b-content-safety
Cutting-edge vision-language model exceling in high-quality reasoning from images.DownloadableFree Endpointllama-3.2-11b-vision-instruct
Cutting-edge vision-Language model exceling in high-quality reasoning from images.DownloadableFree Endpointllama-3.2-90b-vision-instruct
Items per page
of 1 pages