All talks from PyTorch Conference Europe 2026
What PyTorch Conference Europe 2026 Was Really Like – Official PyTorchCon EU Highlights | Paris
This talk captures the vibrant atmosphere of a tech conference focused on artificial intelligence in...
Lightning Talk: How DeepInverse Is Solving Imaging in Science and H... Andrew Wang & Minh Hai Nguyen
This talk covers the significant application of AI in imaging, particularly through the lens of comp...
Why WideEP Inference Needs Data-Parallel-Aware Scheduling - Maroon Ayoub & Tyler Michael Smith
This talk covers the development and functionality of LMD, a distributed large model inference frame...
Write Once, Run Everywhere with Pytorch Transformers - Pedro Cuenca, Hugging Face
This talk discusses how the Transformers library serves as a reference implementation for various do...
Build PyTorch to Understand PyTorch - Vijay Janapa Reddi & Andrea Mattia Garavagno
This talk introduces Tiny Torch, an educational project designed to help learners build their own ma...
Lightning Talk: Why Your Forecasting Transformer Isn’t Working (And How To Fix It... Rosheen Naeem
In this talk, Roshni Naeem discusses challenges in forecasting solar energy generation using a trans...
Lightning Talk: Bayesian Neural Networks With Variational Inference in PyTorch - Lars Heyen
This talk discusses the integration of AI and uncertainty quantification in weather forecasting. The...
Lightning Talk: TerraKit: Standardising AI-Ready Geospatial Data... Rosie Lickorish & Romeo Kienzler
This talk introduces Terra Kit, an open source Python library designed for the curation of geospatia...
Lightning Talk: ExecuTorch on Microcontrollers: Deploying PyTorch To... RJ Ascani & Matthias Cremon
This talk covers Executive Torch, a framework developed by Meta for deploying PyTorch models on micr...
Enabling State-of-the-art Asynchronous Execution in Torch.compile With CUDA Streams - Michael Lazos
This talk discusses the integration of state-of-the-art asynchronous execution in Torch Compile usin...
The Token Slice: Implementing Preemptive Scheduling Via Chunked Decod... Maroon Ayoub & Kellen Swain
This talk discusses the 'token slice' and its role in improving AI workload scheduling through the L...
Lightning Talk: Deep Learning in the Wild: Embedded PyTorch for... Taraqur Rahman & Owen O'Donnell
This talk presents a project focused on deep learning applications in remote environments, specifica...
The Science and Practice of Open and Scalable LLM Evaluations - Grzegorz Chlebus, NVIDIA
This talk focuses on evaluation methodologies at Nvidia, specifically around benchmarking machine le...
Lights, Camera, Inference! Video Generation as a Service With VLLM-O... Ricardo Noriega & Doug Smith
This talk presents VLLM Omni, a powerful inference server designed for multimodal model serving, whi...
Keynote: PyTorch Updates - Edward Yang, Research Engineer, Meta
In this keynote presentation at the first PyTorch conference in Europe, Edward, a core maintainer fr...
Tour De Force: LLM Inference Optimization From Simple To Sophisticated - Christin Pohl, Microsoft
This talk covers LLM inference optimization with a focus on self-hosting models and the consideratio...
Lightning Talk: Bringing Google’s Colossus to PyTorch: Rapid Stor... Ankita Luthra & Trinadh Kotturu
This talk discusses the integration of Google Cloud Storage file system with PyTorch, addressing the...
Helion 1.0: A High-Level DSL for Performance Portable Kernels - Oguz Ulgen, Meta
This talk introduces Helium, a new domain-specific language (DSL) for kernel authoring that simplifi...
torch.compile and Diffusers: A Hands-On Guide to Peak Performance - Sayak Paul, Hugging Face
In this talk, the speaker, Shyuk, a research engineer at Hugging Face, discusses the integration of...
Optimizing PyTorch on CPU-GPU Coherent Platforms - Matthias Jouanneaux, Nvidia
In this session, Matthias, a software engineer for PyTorch at Nvidia, discusses the optimization of...
Lightning Talk: Jigsaw: Domain and Tensor Parallelism for High-Resolution Inp... Deifilia Kieckhefen
This talk presents Jigsaw, a domain and tensor parallel technique aimed at optimizing the training o...
Why Classic IAM Collapses for Agents: Rethinking IAM for Agentic Systems - Parul Singh, Red Hat
This talk examines the limitations of classic identity management (IM) in accommodating modern auton...
Securing Agentic AI With PyTorch: Threat Modeling & LLM Red Teaming in Practice - Valeri Milke
In this session, Sabre Milka discusses the various risks associated with artificial intelligence (AI...
Keynote: Co-Evolution: How the Open Source Intelligence Stack Compounds - Mark Collier
This talk marks the inaugural PyTorch conference in Europe and emphasizes the significance of the Py...
Lightning Talk: Implementing Single-Dim Strategies With Sharding Validator - Anshul Sinha, Meta
This talk focuses on implementing single dimension strategies using the strategy validator in DTenso...
Keynote: Community Led Open Source RL - Joe Spisak
This talk is presented by Joe, a significant contributor to the development of PyTorch, and now a ke...
Lightning Talk: Flexible Deployment of PyTorch Models on MCU-Class... Robert Kalmar & Martin Pavella
In this session, Robert Calmar discusses the flexible deployment of PyTorch models on microcontrolle...
Lightning Talk: TorchJD: Jacobian Descent in PyTorch - Pierre Quinton & Valérian Rey
This talk introduces Torch JD, a library designed for multi-objective optimization using PyTorch, fo...
Lightning Talk: Cross-Region Model Serving: PyTorch Inference, Observability... Suraj Muraleedharan
This talk covers cross-region model serving and multi-region model inference from a customer perspec...
Lightning Talk: Coding Agents for Compiler Construction: Beyond the... Reza Rahimi & Stefan Krassin
This talk features Reza, the CTO of YASSP, as he presents on coding agents for compiler construction...
Optimizing Reinforcement Learning at Trillion-Parameter Scale - Songlin Jiang
This talk covers the application of reinforcement learning on trillion parameter large language mode...
Sponsored Session: TorchTPU: Expanding TPU Programmabil... Kat Ko, Claudio Basile & Jana van Greunen
This talk introduces Torch TPU, a new platform that enhances PyTorch's compatibility with various ha...
Lightning Talk: Training Embedding Model Resiliently for Multimodal M... Huamin Chen & Haichen Zhang
This talk covers the importance and functionalities of a semantic router in model inference for appl...
Lightning Talk: Ethical, Privacy and Sustainability Considerations in PyTorch S... Paula Mesa Macias
In this session, Paula Mesa Matias discusses ethical privacy and sustainability considerations in Py...
Brevitas Quantization Library - Pablo Monteagudo Lago, AMD
This talk covers Brevitas, a neural network quantization framework in PyTorch developed by AMD. The...
Sponsored Keynote: From One Node to Distributed Training and Inference. How the PyTo... Ramine Roane
This talk covers the evolution of PyTorch from simplifying single GPU programming to becoming the pr...
Teaching PyTorch To Read Your Worst PDFs With Docling - Mingxuan Zhao, Peter Staar & Carol Chen
This talk introduces Dog, an open-source Python library developed by IBM that focuses on efficient d...
Lightning Talk: Running ExecuTorch Applications With Silicon Accelera... George Gekov & Aki Makkonen
This talk focuses on running executor applications with silicon acceleration for ultra low power dev...
Lightning Talk: From Pretrained To Personal: Privacy-First... Daniel Holanda Noronha & Iswarya Alex
This talk focuses on fine-tuning machine learning models on AMD's AIPCs, highlighting the benefits o...
Keynote: The Unbearable Lightness of (Agentic) Evaluations - Besmira Nushi
This talk focuses on the evolution of AI evaluation over the last three years, detailing the shift i...
Parameterized CUDA Graph Launch in PyTorch: CUDA Graphs Without the Pain - Daniel Galvez, NVIDIA
In this talk, Daniel Galvez discusses the enhancement of CUDA APIs within PyTorch, focusing specific...
On-Device LLM Inference on Android With ExecuTorch and Qualcomm QNN - Shivay Lamba & Kartikey Rawat
This talk introduces ExecuTorch, an on-device AI framework that enables the running of PyTorch model...
Sponsored Keynote: Any [ Agent | Model | Accelerator | Cloud ]... Maryam Tahhan & Nicolò Lucchesi
This talk focuses on Red Hat's strategy and initiatives in the realm of open-source artificial intel...
Lightning Talk: Combo Kernels: Horizontal Fusion Optimizat... Karthick Panner Selvam & Elias Ellison
This talk discusses combo kernels and horizontal fusion optimization in the context of torch.compile...
How To Write C++ Extensions in 2026 - Jane Xu, Meta & Mikayla Gawarecki, Meta
This talk discusses how to create a C++ custom extension for PyTorch in 2026, emphasizing the proces...
Lightning Talk: Inside VLLM's KV Offloading Connector: Async Memory Transfers for... Nicolò Lucchesi
This talk covers native key-value (KV) cache offloading in the VLM project, which is recognized as t...
Lightning Talk: Accelerating On-Device ML Inference With ExecuTorch and Arm SME2 - Jason Zhu, Arm
This talk presents the collaborative work of Tyler Mullen and Jason Zoo from ARM, focusing on enhanc...
From Responses To Trajectories: Multi-Turn and Multi-Environ... Kashif Rasul & Sergio Paniego Blanco
This talk, presented by Kashif and Sergio from Hugging Face, focuses on the evolution of reinforceme...
Lightning Talk: From Hugging Face To Handheld: Scaling LLM Deployment W... Cormac Brick & Weiyi Wang
This talk focuses on deploying Hugging Face transformer-like language models to edge devices using L...
Beyond the Theory: What Actually Breaks When You Scale Your Disaggregat... Ekin Karabulut & Ron Kahn
This talk discusses the challenges of scaling disaggregated PyTorch models in production environment...
Bringing PyTorch Monarch to AMD GPUs: Single-Controller Distributed Tra... Liz Li & Zachary Streeter
This talk covers the collaboration between AMD and Meta to enable single-controller distributed trai...
Lightning Talk: Why Logging Isn’t Enough: Making PyTorch Training Regressions Vi... Sahana Venkatesh
In this session, Sahana from Wave discusses the challenges of training models at scale, particularly...
Lightning Talk: Graph Based Pipeline Parallelism - Sanket Purandare, Meta & Simon Fan, Meta PyTorch
This talk introduces graph-based pipeline parallelism, a novel approach developed by the Meta PyTorc...
Lightning Talk: Ball Tracking and Detection in Soccer Videos - Comparison of... Maciej Szymkowski
In this lightning talk, Mati Sumkowski from Future Processing discusses his research focused on ball...
Lightning Talk: Backpropagation-Free Optimization in PyTorch - Andrii Krutsylo
This talk discusses gradient-free optimization methods applied in deep learning, particularly implem...
Lightning Talk: Every Millisecond Counts: The Fine-tuning Journey of an Ultra-Eff... Pavel Macenauer
This talk discusses the integration of the PyTorch ecosystem with NXP chips, which range from microc...
Bringing ExecuTorch To the Next Frontiers of Edge AI - Mergen Nachin, Meta
This talk focuses on the future of artificial general intelligence (AGI) and the role of Executor in...
Lightning Talk: Full-Stack PyTorch Robotics VLA: From Data To... Samet Akcay & Dmitriy Pastushenkov
This talk introduces a new full stack robotics vision language action model-based framework develope...
Lightning Talk: Slash LLM Cold-Start Times by Pre-distributing GPU... Billy McFall & Maryam Tahhan
This talk covers strategies for reducing the startup time of large language model (LLM) deployments...
Lightning Talk: Not All Tokens Are Equal: Semantic KV-Cache for Agen... Maroon Ayoub & Hyunkyun Moon
This talk explores the next set of optimizations in KV cache-centric inference, building on concepts...
Lightning Talk: Debugging the Undebuggable: Introducing Torch.distributed.debug - Tristan Rice
This talk presents improvements to debugging in PyTorch distributed, emphasizing the new 'torch dist...
Lightning Talk: KV-Cache Centric Inference: Building a State-Aware... Maroon Ayoub & Martin Hickey
This talk covers the critical concepts of service level objectives (SLOs) in AI inferencing, focusin...
Lightning Talk: FlexAttention + FlashAttention-4: Fast and Flexible - Driss Guessous, Meta
In this talk, TrisKus from Meta presents FlexAttention and FlashAttention V4, a new backend that sig...
Lightning Talk: Beyond Generic Spans: Distributed Tracing for Actio... Sally O'Malley & Greg Pereira
This talk covers the integration of tracing into the LMD framework, a Kubernetes-native distributed...
Lightning Talk: Step-Aligned Telemetry for Distributed PyTorch Training (Time... Abhinav Srivastav
This talk discusses step aligned telemetrics for PyTorch distributed training, focusing on how to id...
Optimizing Large MoE Inference on NVIDIA Blackwell: NVFP4, ADP, and DualPipe Strat... Julien Demouth
This talk explores the optimization of Mixture of Experts (MOE) inference on NVIDIA's Blackwell arch...
Lightning Talk: Live Migration of PyTorch GPU Nodes From Azure To European Clouds - Mike Krom
This talk discusses the process of migrating PyTorch projects to European cloud providers, specifica...
Lightning Talk: Enabling the Audio Modality for Language Models - Eustache Le Bihan, Hugging Face
This talk focuses on how to integrate the audio modality into language models within the open source...
Optimizing CPU LLM Inference in PyTorch: Lessons From VLLM - Crefeda Rodrigues & Fadi Arafeh
This talk focuses on optimizing large model inference on CPU within the PyTorch ecosystem, specifica...
TorchStore: What We Learned Building Distributed Storage Sol... Lucas P, Danielle P, Allen W, Amir A
This talk introduces Torch Store, a solution for efficient weight synchronization in reinforcement l...
Lightning Talk: Scaling Recommendation Systems To 2K GPUs and Beyond - Zain Huda, Meta
This talk covers the scaling of recommendation systems using PyTorch, specifically focusing on techn...
De-mystifying PyTorch for ASICs: When (and Why) To Move Your Development To AI A... Alpha Romer Coma
This talk explores the transition of machine learning development from traditional GPUs to AI accele...
Model-Changing Transforms With Torch.compile - Thomas Viehmann, Lightning AI
In this talk, Tom Vman discusses model changing transforms using Torch Compile, highlighting the ben...
PyTorch Symmetric Memory + NCCL Device APIs: A New Path Towards Multi-GP... Ke Wen & Sylvain Jeaugey
This talk focuses on the integration of PyTorch symmetric memory with Nico device APIs, presenting a...
Lightning Talk: Building a PyTorch‑native VLLM Plugin for IBM Spyre - Thomas Parnell & Thomas Ortner
This talk delves into the integration of the Spire device with the VLLM plugin, which enhances perfo...
PyTorch on RISC-V: From Cross-Compilation To Native CI - Ludovic Henry, Meta
This talk explores the integration of PyTorch with RISC-V, highlighting the challenges and collabora...
Sponsored Keynote: Open Source Infrastructure for the AI Native Era - Jonathan Bryce
This talk discusses the growth of the PyTorch community and its relationship with the Cloud Native C...
Keynote: Gemma 4: Compacting Intelligence for the Edge - Léonard Hussenot
This talk explores the evolution of large language models (LLMs), highlighting the shift from pre-tr...
Keynote: Stream Everything - Moving from Request input to Streaming input - Patrick von Platen
In this talk, Patrick from Mistral discusses an innovative approach to streaming in language models...
Lightning Talk: Pluggable PyTorch LLM Inference Architecture With VLL... Yahav Biran & Maen Suleiman
In this talk, Yahav from the Anapuna ML team discusses the challenges of running machine learning mo...
Lightning Talk: Faster Than SOTA Kernels in Torch.compile With Subgrap... Elias Ellison & Paul Zhang
In this talk, Paul Zheng from the PyTorch team at Meta discusses advancements in subgraph fusions an...
From Gradients To Governance: Making PyTorch Lineage-Aware - Kateryna Romashko & Clodagh Walsh
This talk addresses the critical need for lineage awareness in PyTorch as it transitions into enterp...
Lightning Talk: Torch-Spyre: Compiling To a Multi-core Dataflow Accelerator... D. Grove & O. Tardieu
This talk covers the insights gathered by IBM's team regarding their implementation of PyTorch Induc...
Portable High‑Performance LLM Serving: A Triton Backend for... Burkhard Ringlein & Jan van Lunteren
This talk covers the introduction of a Triton backend for vLLM, which is becoming the industry stand...
Lightning Talk: Bridging the Gap: Engineering Compliant... Muhammad Saqib Hussain & Mohaddisa Maryam
This talk explores the intersection of artificial intelligence and clinical practice, highlighting t...
Keynote: The Hub as Infrastructure. From Open PyTorch Models, to a Safe and Perfor... Lysandre Debut
In this talk, Nander, the chief open source officer at Hugging Face, discusses the development of th...
Seamless Integration: Custom Kernels in the Torch.compile Stack Wi... Kshiteej K, Masaki K & Pawel G
This talk covers the integration of custom kernels into the torch compile stack, focusing on the col...
Keynote: PyTorch CTO - Matt White, Global CTO of AI, Linux Foundation
This talk provides an overview of the current state and growth of the PyTorch Foundation and its eco...
Deploying PyTorch Models To the Browser and Beyond With Transformers.js - Joshua Lochner
In this talk, Joshua, a machine learning engineer at Hugging Face, discusses his journey and the cre...
Lightning Talk: Building AI That Ops Teams Actually Trust - Robert King
In this talk, Rob King discusses building trustworthy AI solutions for operations teams, which often...
Lightning Talk: Distributed AI Without the Infrastructure Tax - Yahav Biran & Maen Suleiman
In this session, Yaav Biran, a customer engineer from the Annapurna ML team at Amazon, addresses the...
Fp8 Training From Hopper To Blackwell - Luca Wehrstedt, Meta
In this talk, the speaker discusses the implementation of FP8 training on Nvidia's latest GPUs, emph...
DualPipe from Scratch: Implementing DeepSeek's 5D Parallelism in PyTorch - Dev Jadhav, ING Bank
This talk focuses on the Dual Pipe implementation in PyTorch, presented by Theo Jadav, a tech lead m...
Building Trust for Users and Regulators Alike: A Cost-Efficient PyTorch Pat... Raja Gopal Hari Vijay
This talk focuses on the integration of compliance mechanisms within the Pyarch framework to address...
Keynote: vLLM & Ray Updates - Tyler Michael Smith & Artur Niederfahrenhorst
This talk covers advancements in vLLM, spearheaded by Tyler Smith, the chief architect of inference...
Beyond JSON-RPC: Scaling Model Context Protocols With gRPC in the Py... Ashesh Vidyut & Madhav Bissa
This talk explores the integration of GRPC with the Model Context Protocol (MCP) within the PyTorch...
Lightning Talk: Achieving SOTA GEMM Performance: A CuTeDSL Backend for PyTorch Induc... Nikhil Patel
This talk covers recent advancements in integrating QTSL gems into PyTorch Inductor to enhance perfo...
Bridging the Hardware Gap With Code Harnesses on the Hugging Face Kernels Hub - Ben Burtenshaw
This talk covers the integration of coding agents with custom kernels to address the growing demand...
Sponsored Session: Fault-Tolerant Training: How We Build Rel... Cyril Konkratenko & Maurits de Groot
This talk addresses the challenges of GPU cluster failures and presents strategies for minimizing su...
Accelerating Complex-Valued Tensors With Torch.compile - Hameer Abbasi, OpenTeams Inc.
This talk addresses the issue of compiling complex tensors in PyTorch, which currently cannot be com...
Lightning Talk: Accelerating PyTorch Models With Torch.compile's C++ Wrapper Mode - Bin Bao, Meta
This talk discusses the acceleration of PyTorch models using the torch.compile C++ wrapper mode pres...
Lightning Talk: Trinity Large - Torchtitan on 2000+ B300s - Matej Sirovatka, Prime Intellect
This talk focuses on the journey of pre-training a 400 billion parameter model from scratch, utilizi...
Lightning Talk: Monarch: An API To Your Supercomputer - Marius Eriksen, Meta
In this talk, Marius Eriksen presents Monarch, a project from Meta that provides a Pythonic API for...