Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
ForumsSupport
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

1 results for

Filters (1)

Use Case
Inference Providers
Publisher
Audience
Blueprint Type
Domain
NIM Container GPUs
Library
Labels (1)
DGX Station
30 MIN

LLM Inference with SGLang

Serve LLMs with SGLang on DGX Station (Qwen3-8B default; Qwen3.6 MoE optional)—prefix-cached multi-turn, structured output, benchmarks, and inference-server guidance
Playbook
  • RadixAttention
  • Structured Output
  • Blackwell
  • DGX Station
  • Inference
  • SGLang
  • GB300
Last updated on May 26, 2026
Items per page
of 1 pages