Keynote: Gemma 4: Compacting Intelligence for the Edge - Léonard Hussenot
About this talk
This talk explores the philosophy and engineering behind Gemma 4, emphasizing that the future of AI depends on "intelligence per byte" rather than just size. The speaker discusses the importance of compacting intelligence, which maximizes the reasoning and instruction-following ability of each token, as a crucial factor for developing genuinely useful AI. By focusing on token efficiency and minimizing memory footprints, this approach enables the creation of faster, more private, and accessible applications in the field of artificial intelligence.
More from this event
See all 103 talks →
What PyTorch Conference Europe 2026 Was Really Like – Official PyTorchCon EU Highlights | Paris
0:53
Lightning Talk: How DeepInverse Is Solving Imaging in Science and H... Andrew Wang & Minh Hai Nguyen
9:50
Why WideEP Inference Needs Data-Parallel-Aware Scheduling - Maroon Ayoub & Tyler Michael Smith
25:37
Write Once, Run Everywhere with Pytorch Transformers - Pedro Cuenca, Hugging Face
19:17