Bridging the Hardware Gap With Code Harnesses on the Hugging Face Kernels Hub - Ben Burtenshaw
About this talk
This talk covers the development of code harnesses for standardizing kernel writing in the context of agentic coding. The speaker, Ben Burtenshaw from Hugging Face, presents an end-to-end experiment that benchmarks six harnesses across ten models using CUDA and Metal. Key topics include agent cost, kernel latency, VRAM usage, and end inference performance, highlighting how the Kernels Hub facilitates large-scale distribution. The talk also introduces two tools: the Kernels Hub, an infrastructure for creating reproducible kernels in the PyTorch ecosystem, and HF Skills, a library for defining and evaluating skills necessary for machine learning tasks related to kernel writing. This session is aimed at kernel writers and PyTorch developers seeking to enhance their understanding of high-performance kernel optimization.
More from this event
See all 103 talks →
What PyTorch Conference Europe 2026 Was Really Like – Official PyTorchCon EU Highlights | Paris
0:53
Lightning Talk: How DeepInverse Is Solving Imaging in Science and H... Andrew Wang & Minh Hai Nguyen
9:50
Why WideEP Inference Needs Data-Parallel-Aware Scheduling - Maroon Ayoub & Tyler Michael Smith
25:37
Write Once, Run Everywhere with Pytorch Transformers - Pedro Cuenca, Hugging Face
19:17