What is an AI Engineer at AMD?
At AMD, the AI Engineer role operates at the cutting edge of modern high-performance computing, hardware acceleration, and generative artificial intelligence. As frontier models—such as Large Language Models (LLMs), Vision-Language Models (VLMs), and Mixture-of-Experts (MoE) architectures—scale exponentially in parameter count and complexity, the underlying compute stack requires relentless optimization. The AMD AI Group builds the critical bridge between state-of-the-art machine learning algorithms and high-throughput hardware architectures, ensuring that AMD Instinct GPUs and the ROCm ecosystem deliver premier performance across data centers, supercomputers, and cloud environments.
As an AI Engineer, your work directly impacts how industry-leading organizations, researchers, and enterprise customers train and serve frontier models. You will be responsible for designing end-to-end model execution frameworks, writing performance-critical GPU kernels in HIP or Triton, and integrating optimized runtimes into core open-source serving engines such as vLLM and SGLang. Your contributions ensure that AMD platforms serve as first-class hardware targets for distributed LLM serving, retrieval-augmented generation (RAG) platforms, agentic workflows, and large-scale AI cluster infrastructure.
This role requires a rare synthesis of deep systems engineering and advanced machine learning expertise. Whether you are tuning collective communication patterns with RCCL, developing custom quantization schemes (such as FP8, FP4, or AWQ), or architecting disaggregated LLM serving systems, you will operate in a dynamic, high-impact environment where software performance unlocks the full capability of physical silicon.

