1. What is a Site Reliability Engineer at Arista Networks?
As a Site Reliability Engineer (SRE) within the Engineering Productivity (EngProd) team, you are the backbone of Arista Networks’ innovation engine. You are responsible for the systems and infrastructure that empower over 2,000 engineers to build the routing and switching products that define the industry’s largest data centers. In this role, you aren’t just keeping the lights on; you are actively architecting, scaling, and automating the environments that handle massive scale—including terabytes of source control, hundreds of thousands of daily build jobs, and complex Kubernetes clusters.
This position is critical because you own the developer experience. When the infrastructure you manage performs reliably and efficiently, the entire engineering organization moves faster. You will face challenges involving high-concurrency systems, complex CI/CD pipelines, and the need for deep technical intuition. If you are passionate about debugging at scale, automating away toil, and building tools that improve the daily lives of software engineers, this role offers a unique opportunity to influence the operational excellence of a market-leading technology company.

