NVIDIA · NCP-GENL

The compute-efficiency thread

Getting more accuracy per unit of compute. Opens in M1 with training-stability levers (normalization, LR warmup, gradient clipping), and concentrates in M5's mixed precision, quantization, and pruning before M6 puts TensorRT and Triton underneath the served pipeline.

NCAM-T3 · 0 lessons across 0 modules

    Part of the throughlines running across the NCP-GENL prep course.