Theme 1 – Feasibility and limits of running CUDA on AMD hardware
Many commenters discuss the promise (and current limits) of projects like ZLUDA that let CUDA‑targeted code run on AMD GPUs.
“RDNA1 isn't good for a whole lot, even flagship RDNA2 cards are a stretch for many things. The lack of WMMA/matrix multiply/BF16 is too severe of a penalty.” – monster_truck
“I wish there were a way to use RDNA1 cards with CUDA for AMD. My 5700XTs are sitting in a drawer.” – system2
Theme 2 – Nvidia’s CUDA moat and ecosystem inertia
The discussion repeatedly points to CUDA’s entrenched advantage, suggesting that even if translation becomes easy, the surrounding stack keeps Nvidia ahead.
“Nvidia's moat is (and will remain for the foreseeable future) the entire stack. you cannot fathom the pain and misery of working on literally any other stack.” – mathisfun123
“AI will take down Nvidia’s moat. When it becomes trivial to translate CUDA/PTX to HIP, SYCL or Metal, CUDA is no longer the moat, it becomes the intermediate representation.” – swerner
Theme 2 – Need for and frustrations with open standards
Several users argue that a unified, vendor‑neutral approach (HIP, SYCL, OpenCL, OneAPI) is essential but hampered by fragmentation and poor developer experience.
“Off-topic and somewhat of a rant, but I'd far prefer us all focusing on open standards like HIP, SYCL, OpenCL, etc.” – linuxhansl
“I would too, but sadly that's Khronos' job to organize, and they've had trouble getting American vendors to work together.” – bigyabai
“The problem with OneAPI is naming. It leads people to believe that is another competing standard where in fact is is simply just an implementation of a standard compliant SYCL compiler.” – swerner