2023 LLVM Developers' Meeting
https://llvm.org/devmtg/2023-10
------
Optimization of CUDA GPU Kernels and Translation to AMDGPU in 4) Polygeist/MLIR
Ivan Ivanov
------
Slides: Coming Soon
-----
We extend the Polygeist C/C++ compiler to utilize a target-agnostic parallel representation of GPU kernels in MLIR to perform parallel optimizations and architecture-specific tuning. We also implement translation from CUDA to AMDGPU and expand the set of possible target hardware for CUDA code.
-----
Videos Edited by Bash Films: http://www.BashFilms.com