AMD, in collaboration with GPU_MODE, has launched a $1.1 million kernel hackathon that has significantly improved the performance of its MI355X graphics card. The Readonflow Team's optimizations, focusing on MoE kernels and other core components, have resulted in up to a 4x performance increase and have been upstreamed to AMD's AITER kernel library and the ATOM inference engine. This initiative aims to bring AMD's vLLM performance closer to parity with CUDA vLLM and enhance community experience with the ROCm stack. AI
IMPACT Enhances AMD's competitiveness in AI hardware by improving inference performance and fostering community development on its ROCm stack.
RANK_REASON Community-driven performance optimization and upstreaming of kernels for AMD hardware.
- AMD
- Anthropic
- Instella-MoE
- MI300X
- MI325X
- Mythos
- NVIDIA
- ROCm
- ROCmFPX
- White House
- AIatAMD
- AnushElangovan
- ATOM
- marksaroufim
- MI355X
- Nvidia B200
- Readonflow Team
- vLLM
AI-generated summary · Google Gemini · from 6 sources. How we write summaries →