Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
models
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and moreDownloadableFree Endpointnemotron-3-ultra-550b-a55b
Items per page
of 1 pages