Breaking the Monolith: Decomposing and Governing Giant LLM Jobs Across Clusters - Kevin Wang, Huawei
About this talk
This session features Kevin Wang from Huawei discussing the challenges and solutions for managing large language model (LLM) jobs across multiple clusters. The talk introduces the concepts of multi-cluster architecture, with a focus on adaptive cross-cluster scheduling using Volcano Global and Karmada. Key topics include the implementation of a universal global scheduling control plane, the creation of a higher-level job abstraction for intelligent job decomposition, and the establishment of a centralized global queue. This approach aims to enhance resource utilization and flexibility while ensuring fair allocation of resources and preventing bottlenecks in shared environments.
More from this event
See all 436 talks →
Best of KubeCon + CloudNativeCon Amsterdam 2026
2:17
The Quiet Work of Forever: Sustaining Open Source Communities - O. Hope Amaechi-Okorie, JSON Schema
26:24
Evolving KServe: The Unified Model Inference Platform for Both Predictive and... F. Spolti & J. Lee
32:40
Preventing S3 Cost Storms: Applying Cortex’s Efficiency Lessons to I/O-Heav... A. Fishman-Lichterman
5:32