Write Once, Run Everywhere with Pytorch Transformers - Pedro Cuenca, Hugging Face
About this talk
This talk covers the integration of the Hugging Face transformers library, built on pure PyTorch, into various machine learning architectures. The speaker, Pedro Cuenca from Hugging Face, highlights how model definitions serve as reference implementations for training libraries and deployment engines like vLLM and SGLang. The session details the journey towards simplifying the downstream integration of transformers models, enabling seamless interoperability without code modifications. The discussion also touches on future interoperability features with libraries such as MLX and llama.cpp.
More from this event
See all 103 talks →
What PyTorch Conference Europe 2026 Was Really Like – Official PyTorchCon EU Highlights | Paris
0:53
Lightning Talk: How DeepInverse Is Solving Imaging in Science and H... Andrew Wang & Minh Hai Nguyen
9:50
Why WideEP Inference Needs Data-Parallel-Aware Scheduling - Maroon Ayoub & Tyler Michael Smith
25:37
Build PyTorch to Understand PyTorch - Vijay Janapa Reddi & Andrea Mattia Garavagno
21:08