<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>AMD on moyutianzun 的博客</title><link>https://moyutianzun.com/tags/amd/</link><description>Recent content in AMD on moyutianzun 的博客</description><generator>Hugo</generator><language>zh-cn</language><copyright>moyutianzun</copyright><lastBuildDate>Fri, 19 Sep 2025 02:57:25 +0800</lastBuildDate><atom:link href="https://moyutianzun.com/tags/amd/index.xml" rel="self" type="application/rss+xml"/><item><title>AMD 2025 分布式推理算子优化挑战赛 —— lect 9/16 note</title><link>https://moyutianzun.com/blog/amd-2025-fen-bu-shi-tui-li-suan-zi-you-hua-tiao-zhan-sai------lect-9-16-note/</link><pubDate>Fri, 19 Sep 2025 02:57:25 +0800</pubDate><guid>https://moyutianzun.com/blog/amd-2025-fen-bu-shi-tui-li-suan-zi-you-hua-tiao-zhan-sai------lect-9-16-note/</guid><description>&lt;p style=""&gt;&lt;/p&gt;&lt;h1 style="" id="rocm-%E5%85%A5%E9%97%A8"&gt;ROCm 入门&lt;/h1&gt;&lt;p style=""&gt;&lt;img src="https://moyublog-picture.oss-cn-guangzhou.aliyuncs.com/images/20250916190914403.png" width="100%" height="100%" style="display: inline-block"&gt;&lt;/p&gt;&lt;p style="text-indent: 2em"&gt;首先就是amd官方的命名跟nv的区别，其实区别并不大，只是AMD在cuda的基础上做了更多的优化，比如说一个wavefront有64个work-item，相当于一个warp有64个threads。其次就是有两种register，在&lt;/p&gt;</description></item><item><title>AMD 2025 分布式推理算子优化挑战赛——笔记</title><link>https://moyutianzun.com/blog/amd-2025-fen-bu-shi-tui-li-suan-zi-you-hua-tiao-zhan-sai----bi-ji/</link><pubDate>Tue, 09 Sep 2025 05:48:15 +0800</pubDate><guid>https://moyutianzun.com/blog/amd-2025-fen-bu-shi-tui-li-suan-zi-you-hua-tiao-zhan-sai----bi-ji/</guid><description>&lt;p style=""&gt;比赛提供的link：&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://modelscope.cn/competition/117/%E6%AF%94%E8%B5%9B%E7%AE%80%E4%BB%8B" target="_self" rel=""&gt;魔搭社区比赛首页&lt;/a&gt; &lt;a href="https://www.datamonsters.com/amd-developer-challenge-2025" target="_self" rel=""&gt;AMD比赛首页&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://www.gpumode.com/v2/leaderboard/563?tab=rankings" target="_self" rel=""&gt;amd-all2all kernel Leaderboard&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://github.com/gpu-mode/reference-kernels/tree/main/problems/amd_distributed" target="_self" rel=""&gt;reference-kernels&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://discord.com/channels/" target="_self" rel=""&gt;discord link&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://github.com/gpu-mode/popcorn-cli?tab=readme-ov-file" target="_self" rel=""&gt;Popcorn CLI&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;lect：&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://www.youtube.com/watch?v=dNWv3qYU60E" target="_self" rel=""&gt;ytb Bonus Lecture: AMD Developer Challenge&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://stormy-sailor-96a.notion.site/Mixture-of-Experts-AMD-Problem-1d7221cc2ffa80f9b171c332aed16093" target="_self" rel=""&gt;Mixture of Experts AMD Problem&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;a href="https://moyutianzun.cn/archives/amd-2025-fen-bu-shi-tui-li-suan-zi-you-hua-tiao-zhan-sai------lect-9-16-note" target="_self" rel=""&gt;9/16 lect note&lt;/a&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;&lt;p style=""&gt;&lt;/p&gt;</description></item></channel></rss>