As an AI Engineer at Amazon Web Services, you will operate at the cutting edge of cloud-scale artificial intelligence and machine learning infrastructure. This role sits at the intersection of applied machine learning research and high-performance systems engineering, where you will build, scale, and optimize architectures that power massive generative AI workloads, custom silicon integration (such as AWS Trainium and Inferentia), and distributed training frameworks. You are not just building software; you are architecting the foundational compute layers that enable global enterprises and AI developers to push the boundaries of what is possible with large language models, multimodal models, and advanced agentic systems.
The impact of this position is profound, directly influencing the performance, cost-efficiency, and reliability of AI applications serving millions of users worldwide. Whether you are optimizing distributed training via FSDP and torchtitan, designing low-latency inference serving stacks with vLLM, or engineering robust Retrieval-Augmented Generation pipelines and multi-agent workflows for enterprise customers, your work shapes the industry standard for cloud-based AI. Expect an environment of high autonomy and rapid innovation where deep technical rigor meets customer-obsessed problem-solving.




