1. What is a Site Reliability Engineer at Cerebras?
As a Site Reliability Engineer at Cerebras, you are at the intersection of breakthrough hardware and cutting-edge AI software. Cerebras is redefining the compute landscape with its wafer-scale architecture, providing the AI power of dozens of GPUs on a single chip. In this role, you aren’t just maintaining servers; you are ensuring the reliability of the world’s fastest AI inference services for partners like OpenAI and other frontier AI labs.
The work is high-stakes and high-impact. Because Cerebras technology delivers 10x the speed of traditional cloud hyperscalers, the SRE function is critical to enabling real-time generative AI applications. Whether you are focusing on platform automation, capacity provisioning, or building declarative GitOps pipelines, your work directly impacts the ability of the world’s leading researchers to deploy and iterate on massive models. You will be part of a team building the "tomorrow" layer of AI infrastructure, where engineering rigor and operational excellence are the primary drivers of success.




