NVIDIA · NCP-GENL
The compute-efficiency thread
Getting more accuracy per unit of compute. Opens in M1 with training-stability levers (normalization, LR warmup, gradient clipping), and concentrates in M5's mixed precision, quantization, and pruning before M6 puts TensorRT and Triton underneath the served pipeline.
NCAM-T3 · 0 lessons across 0 modules
Part of the throughlines running across the NCP-GENL prep course.