Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    18 results for

    Filters

    Use Case
    Inference Providers
    Publisher
    Audience
    Blueprint Type
    Domain
    Library

    Search results

    • DGX Spark
      20 MINS

      CLI Coding Agent

      Build local CLI coding agents with Ollama
      Playbook
      • Coding
      • Ollama
      • Claude Code
      • OpenCode
      • Qwen
      • LLM
      • Codex
      Last updated on April 27, 2026
    Items per page
    of 1 pages
    DGX Station
    30 MINS

    Local Coding Agent

    Run local CLI coding agents with Claude Code and Ollama on DGX Station (NVIDIA GB300) using qwen3.6:27b
    Playbook
    • Coding
    • Ollama
    • Claude Code
    • DGX Station
    • LLM
    • GB300
    Last updated on March 25, 2026
  • Use this skill when the user is doing hands-on DOCA Erasure Coding programming on a BlueField DPU, ConnectX NIC, or host — bringing up a doca_ec context, picking among the create / recover / update tasks, choosing matrix type / N / K / block size, queryin
    Skill
    • Developer
    • Platform Engineer
    • Hands On Builder
    • Application Developer
    • Accelerated Computing
    • DOCA
    50 downloads in the last 30 days
    Last updated on July 20, 2026
  • DGX Spark
    30 MIN

    Vibe Coding in VS Code

    Use DGX Spark as a local or remote Vibe Coding assistant with Ollama and Continue
    Playbook
    • DGX
    • VibeCoding
    • Spark
    Last updated on October 10, 2025
  • DeepSeek AI
    Free Endpoint

    deepseek-v4-pro-0813

    DeepSeek V4 scales to 1M-token context windows with efficient MoE architecture for coding tasks.
    Model
    • coding
    • Moe
    • reasoning
    • agentic
    Last updated on August 26, 2026
  • Google
    DownloadableFree Endpoint

    gemma-4-31b-it

    Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.
    Model
    • reasoning
    • coding
    • text-to-text
    • agentic
    6M API calls in the last 30 days
    Last updated on April 2, 2026
  • Poolside
    Free Endpoint

    laguna-xs-2.1

    Efficient 33B MoE for local, long-horizon agentic coding and terminal tasks
    Model
    • Agentic AI
    • Coding
    • Reasoning
    • Tool Use
    Last updated on July 15, 2026
  • Minimaxai
    Deprecation in 11dFree Endpoint

    minimax-m3

    MiniMax M3 Preview is a multimodal MoE vision-language model with strong reasoning, coding, and tool-calling capabilities.
    Model
    • coding
    • text-to-text
    • reasoning
    10M API calls in the last 30 days
    Last updated on June 12, 2026
  • NemoClaw
    NemoClaw

    NemoClaw for LangChain Deep Agents Code

    Run open-source Deep Agents Code, tuned for Nemotron 3 Ultra, to plan, edit and test code with enterprise governance.
    Blueprint
    • NVIDIA AI
    • security
    • langchain
    • sandbox
    • deep agents
    • coding agent
    Last updated on July 8, 2026
  • Moonshotai
    DownloadableFree Endpoint

    kimi-k3

    ~2.8T hybrid KDA+MLA multimodal MoE for long-horizon coding, agentic tool use, and image understanding.
    Model
    • Multimodal
    • Mixture-of-Experts
    • Reasoning
    • Image-to-Text
    Last updated on August 27, 2026
  • Mistral AI
    Free Endpoint

    mistral-nemotron

    Built for agentic workflows, this model excels in coding, instruction following, and function calling
    Model
    • language generation
    • instruction following
    • function calling
    1M API calls in the last 30 days
    Last updated on June 12, 2025
  • DeepSeek AI
    Free Endpoint

    deepseek-v4-flash-0731

    284B MoE (13B active) model ideal for long-context workloads optimized for coding, chat, and agentic workflows
    Model
    • MoE
    • Reasoning
    • Long Context
    • Hybrid Attention
    Last updated on August 19, 2026
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-nano-30b-a3b

    Open, efficient MoE model with 1M context, excelling in coding, reasoning, instruction following, tool calling, and more
    Model
    • MoE
    • Reasoning
    • Long Context
    • Instruction Following
    12M API calls in the last 30 days
    Last updated on December 15, 2025
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-super-120b-a12b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    • MoE
    • Reasoning
    • Chat
    • Long Context
    • Instruction Following
    65M API calls in the last 30 days
    Last updated on March 11, 2026
  • NVIDIA
    DownloadableFree Endpoint

    nemotron-3-ultra-550b-a55b

    Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
    Model
    • Agent
    • MoE
    • Frontier
    • Reasoning
    • Long Context
    52M API calls in the last 30 days
    Last updated on June 4, 2026
  • Manage durable working-session memory for coding agents. Use when a user asks to preserve or recover agent context across disconnects, VS Code restarts, long-running work, handoffs, or any session where important state should be written periodically under
    Skill
    • Developer
    • AI Engineer
    • Ml Engineer
    • NeMo RL
    • Developer Tools
    2K downloads in the last 30 days
    Last updated on June 5, 2026
  • Guides human users' AI agents to the NemoClaw docs MCP server and canonical Fern documentation in Markdown form. Use when users ask how to install, configure, operate, troubleshoot, secure, or learn NemoClaw with an AI coding assistant. Trigger keywords -
    Skill
    • Developer
    • AI Engineer
    • DevOps Engineer
    • Platform Engineer
    • Application Developer
    • NemoClaw
    • Developer Tools
    876 downloads in the last 30 days
    Last updated on June 26, 2026
  • General
    Developer Example

    Nsight Copilot - AI Code Assistant for CUDA Development

    Deploy an AI-powered coding assistant on DGX Spark that delivers expert CUDA-aware chat, real-time code completion, and retrieval-augmented generation grounded in authoritative GPU programming knowledge—powered by NVIDIA NIM microservices.
    Blueprint
      Last updated on June 10, 2026