Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
models
Efficient 33B MoE for local, long-horizon agentic coding and terminal tasksFree Endpointlaguna-xs-2.1
Dense 31B model delivering frontier reasoning for coding, agentic workflows, and fine-tuning.DownloadableFree Endpointgemma-4-31b-it
Items per page
of 1 pages