Project Lightning Talk: Hami: Dynamic, Smart, Stable GPU-Sharing Middleware In Kubern... Mengxuan Li
About this talk
This talk covers HAMi, a dynamic middleware designed for efficient GPU-sharing in Kubernetes environments, presented by Mengxuan Li, a maintainer of the project. In the context of AI and machine learning, the speaker discusses the challenges of managing diverse GPU workloads across various hardware vendors and how HAMi addresses key issues such as exclusive allocation, device topology awareness, and observability. The session highlights HAMi's capabilities, including safe device sharing among containers, scheduling optimization through inter-topology analysis, and a unified monitoring system for heterogeneous devices. With support for over 10 hardware vendors and adoption by more than 100 enterprises globally, HAMi has been instrumental in improving GPU utilization across extensive deployments.
More from this event
See all 436 talks →
Best of KubeCon + CloudNativeCon Amsterdam 2026
2:17
The Quiet Work of Forever: Sustaining Open Source Communities - O. Hope Amaechi-Okorie, JSON Schema
26:24
Evolving KServe: The Unified Model Inference Platform for Both Predictive and... F. Spolti & J. Lee
32:40
Preventing S3 Cost Storms: Applying Cortex’s Efficiency Lessons to I/O-Heav... A. Fishman-Lichterman
5:32