
Extending TorchInductor with FlyDSL: A New MLIR-Native Backend for High-Performance GEMMs
Liz Li and Jiahui Cao of AMD will present “Extending TorchInductor with FlyDSL: A New MLIR-Native Backend for High-Performance GEMMs” at PyTorch Conference North America 2026.
The session will cover how FlyDSL, AMD’s Python-native and MLIR-based GPU kernel DSL, integrates with TorchInductor’s GEMM compilation and autotuning pipeline while preserving the torch.compile user experience. Liz and Jiahui will also present performance results for transformer training and inference on AMD Instinct GPUs, comparing Triton and FlyDSL implementations.
Register for PyTorch Conference North America 2026: https://hubs.la/Q04w5M9L0
#PyTorchCon
The session will cover how FlyDSL, AMD’s Python-native and MLIR-based GPU kernel DSL, integrates with TorchInductor’s GEMM compilation and autotuning pipeline while preserving the torch.compile user experience. Liz and Jiahui will also present performance results for transformer training and inference on AMD Instinct GPUs, comparing Triton and FlyDSL implementations.
Register for PyTorch Conference North America 2026: https://hubs.la/Q04w5M9L0
#PyTorchCon
PyTorch
Welcome to the official PyTorch YouTube Channel. Learn about the latest PyTorch tutorials, new, and more.
PyTorch is an open source machine learning framework that is used by both researchers and developers to build, train, and deploy ML systems that so...
torch.compile and Diffusers: A Hands-On Guide to Peak Performance - Sayak Paul, Hugging Face
PyTorch
torch compile and Diffusers - A Hands On Guide to Peak Performance - PyTorch Compiler Series
PyTorch
verl: An Open Source Large Scale LLM RL Framework for Agentic Tasks - Yuxuan Tong, Bytedance
PyTorch