
Maximize GPU Efficiency with Continuous Profiling for GPUs
Matthias Loibl of Polar Signals explains continuous GPU profiling that combines NVIDIA NVML utilization, memory, power, temperature, clock-speed, and PCIe metrics with low-overhead Linux eBPF CPU profiles. He demonstrates correlating GPU underutilization with Python and CUDA call stacks in flame charts, measuring CUDA kernel execution time, and deploying…
Matthias Loibl
RAG, context, and search · Infrastructure and deployment · Evals