Lisa tweet: https://x.com/LisaSu/status/1706707561809105331?s=20
Lamini tweet: https://x.com/realSharonZhou/status/1706701693684154766?s=20
Blog: https://www.lamini.ai/blog/lamini-amd-paving-the-road-to-gpu...
Register: https://www.theregister.com/2023/09/26/amd_instinct_ai_lamin... CRN: https://www.crn.com/news/components-peripherals/llm-startup-...
The hard part about using any AI Chips other than NVIDIA has been software. ROCm is finally at the point where it can train and deploy LLMs like Llama 2 in production.
If you want to try this out, one big issue is that software support is hugely different on Instinct vs Radeon. I think AMD will fix this eventually, but today you need to use Instinct.
We will post more information explaining how this works in the next few weeks.
The middle section of the blog post above includes some details including GEMM/memcpy performance, and some of the software layers that we needed to write to run on AMD.