AI Agents for LLM Inference Runtimes on Edge Hardware
At PyTorch Conference North America, Thomas Cottenier from Arm will present how AI agents can synthesize customized, target-specific runtimes using PyTorch components like torch.export, ExecuTorch, and torchao.
This approach removes redundant prefill computation and creates streamlined inference pipelines that deliver high performance directly on edge hardware.

Join us San Jose this October 20-21 to learn more: https://hubs.la/Q04v4SL60
PyTorch
Welcome to the official PyTorch YouTube Channel. Learn about the latest PyTorch tutorials, new, and more. PyTorch is an open source machine learning framework that is used by both researchers and developers to build, train, and deploy ML systems that so...