Models
Deploy and scale models on your GPU infrastructure of choice with NVIDIA NIM inference microservices
models
Diffusion-based 26B parameter LLM enabling parallel token generation for real-time text appsDownloadableFree Endpointdiffusiongemma-26b-a4b-it
Items per page
of 1 pages