Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
Help Center
Getting Started
  1. Create and verify your account to unlock full access to NVIDIA NIM APIs.
ResourcesDeveloper ForumsContact Support
FAQs
  • Terms of Use
    Privacy Policy
    Your Privacy Choices
    Contact

    Copyright © 2026 NVIDIA Corporation

    1 results for

    Filters (1)

    Publisher
    Labels (1)

    Search results

    • Playbooks
      60 MIN

      Quantize Models to NVFP4 with NVIDIA Model Optimizer

      Cut memory ~3.5× vs FP16 while keeping accuracy close to FP8, then validate with an OpenAI-compatible endpoint
      Playbook
      • vLLM
      • DGX Spark
      • Model Optimizer
      • DGX Station
      • TensorRT-LLM
      • Inference
      Last updated on August 4, 2026
    Items per page
    of 1 pages