You are currently viewing SemiWiki as a guest which gives you limited access to the site. To view blog comments and experience other SemiWiki features you must be a registered member. Registration is fast, simple, and absolutely free so please, join our community today!
For more than a decade, AMD has been chasing Nvidia in AI—but CUDA has remained the industry's biggest barrier. That dynamic may finally be shifting.
This story explores why agentic AI could weaken CUDA's long-standing software moat, why AMD may be better positioned than ever to capitalize, and what this turning point could mean for the broader AI hardware ecosystem.
If CUDA's advantage is no longer untouchable, AMD's biggest opportunity may have finally arrived.
A transcript of a four-hour recording from DeepSeek founder Liang Wenfeng’s(梁文鋒) first fundraising meeting in May, held as the company prepared for a future IPO, has been leaked.
Bloomberg reported a few days ago that Liang was so angered by the leak that he suspended the second round of fundraising. To me, that effectively confirms the authenticity of the leaked conversation.
The remarkably candid private remarks from arguably the most important figure in China’s AI industry today contain a wealth of valuable information. Here are a few highlights.
From my perspective, CUDA hasn't been the hardware moat for data center inference for at least the past year. With open models and open-sourced model serving engines like vLLM and SGlang, the real moat is the set of cost/power/throughput/latency Pareto curves for systems for a broad range of models. And the moat behind that is the speed and capability of each hardware supplier for system-level co-optimization for large MoE attention-based models. One great example here:
Agentic Kernel Generation, Improvement in Software Quality, Unstable Internal Development Clusters, Helios MI455X Production Ramp Hell, Up to 105% Discounts from Finance Engineering
newsletter.semianalysis.com
ps: thought the most important insight from Liang's transcript was the 4x lower performance plus "2 years later" penalty of Chinese / Huawei data center systems.