Skip to main content
NVIDIA
Explore
Models
Skills
Blueprints
GPUs
Docs
ForumsSupport
Terms of Use
Privacy Policy
Your Privacy Choices
Contact

Copyright © 2026 NVIDIA Corporation

1 results for

Filters (1)

Use Case
Inference Providers
Publisher
Audience
Blueprint Type
Domain
NIM Container GPUs
Library
Labels (1)
DGX Station
2 HRS

Profiler-Driven Kernel Optimization for Fine-Tuning

Use torch.profiler to find training bottlenecks, then write custom Triton kernels to optimize LLaMA 8B fine-tuning
Playbook
Training
Fine-TuningPerformance OptimizationKernel DevelopmentDGX StationLLaMATritonGB300
Last updated on May 26, 2026
Items per page
of 1 pages