Search results
A self-hosted browser interface with models running locally on your GPUPlaybooksIntermediate15 MINChat with LLMs Using Open WebUI and Ollama
Natural and expressive voices in 23 languages. For voice agents and brand ambassadors.Downloadablechatterbox-multilingual-tts
Deploy a multi-agent chatbot system and chat with agents on your Spark Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&AGeneralLaunchableEnterpriseBuild a Video Search and Summarization (VSS) Agent
Smaller Mixture of Experts (MoE) text-only LLM for efficient AI reasoning and mathDownloadableFree Endpointgpt-oss-20b
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and moreDownloadableFree Endpointnemotron-3-super-120b-a12b
Multimodal 320B-total / 18B-active MoE with hybrid KDA and sparse MLA attention, native FP8 weights, reasoning and tool calling.Free Endpointglm-5-3-flash
Muse Glimmer 30B is a multimodal reasoning model accepting text and images, with native tool-calling and separate reasoning output.DownloadableFree Endpointmuse-glimmer-30b
Chat from the terminal against local vLLM with the self-improving Nous Research agent (Telegram optional) Model for object detection, fine-tuned to detect charts, tables, and titles in documents.Downloadablenemotron-graphic-elements-v1
Model for object detection, fine-tuned to detect charts, tables, and titles in documents.Downloadablenemotron-page-elements-v3
Model for object detection, fine-tuned to detect charts, tables, and titles in documents.Downloadablenemotron-table-structure-v1
Deploy an AI-powered coding assistant on DGX Spark that delivers expert CUDA-aware chat, real-time code completion, and retrieval-augmented generation grounded in authoritative GPU programming knowledge—powered by NVIDIA NIM microservices.GeneralDeveloper ExampleNsight Copilot - AI Code Assistant for CUDA Development
Use this skill when deploying standalone RT-VLM dense captioning or calling its REST API (uploads, captions, streams, chat-completions, Kafka). Not for VSS profile deploy or video-search ingestion. Automatic speech recognition model that transcribes speech in lower case Spanish with record-setting accuracy and performanceDownloadableconformer-ctc-asr
Set up a cluster of DGX Spark devices that are connected through Switch Vision language model that excels in understanding the physical world using structured reasoning on videos or images.DownloadableFree Endpointcosmos3-nano-reasoner
Use this skill when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether operators perform assembly-line steps in order via event boundary d Use this skill when the user is deploying or operating the DOCA Argus Service — the packaged BlueField-side runtime-security container that watches the BlueField and attached host for suspicious activity, integrity violations, and operational anomalies, a Use this skill when the operator is authoring, building, loading, or debugging a custom doca-bench plug-in — a versioned shared library with DOCA_EXPERIMENTAL-marked C entry points that doca-bench loads to measure a workload class its built-in modes do no WARNING: guides potentially IRREVERSIBLE BlueField-4 hardware operations (PLDM firmware burns, ISO reflashes, power cycles, BMC factory resets) that can brick firmware, corrupt boot media, or cause outages — a maintenance window and rollback plan are requ Use this skill when the user wants to invoke the read-only doca_caps CLI to ask what DOCA sees on this host — listing DOCA devices and PCIe addresses, listing representor devices, asking which DOCA libraries are available on the current OS, checking per-d
Items per page
of 3 pages