Day 0 Support for Gemma 4 on AMD Processors and GPUs
Summary
Gemma 4 is Google’s family of open-weight multimodal models ranging from 2B to 31B parameters, with MoE and dense variants and support for up to 256K token contexts across text, vision, and select audio tasks. AMD announced Day Zero support for the full Gemma 4 lineup across its AI-enabled hardware, including Instinct datacenter GPUs, Radeon workstation GPUs, and Ryzen AI processors for PCs. AMD integrated Gemma 4 with popular AI tools and open-source frameworks such as LM Studio, vLLM, SGLang, llama.cpp, Ollama, and Lemonade to enable optimized inference and multi-request handling. You can deploy Gemma 4 on AMD GPUs using vLLM via Docker or Python installs, with support planned in upstream and nightly builds to target agentic AI workflows and local hardware deployments.