What is a Site Reliability Engineer at Cohesity?
As a Site Reliability Engineer (SRE) at Cohesity, you sit at the intersection of software engineering and systems operations. You are responsible for ensuring the availability, scalability, and performance of Cohesity’s data management and protection platforms. In an environment where data integrity and uptime are mission-critical for enterprise customers, your work directly influences the reliability of the infrastructure that protects their most valuable assets.
This role is not just about keeping the lights on; it is about building automated, resilient systems that can handle massive scale and complexity. You will collaborate closely with engineering teams to bridge the gap between development and production, focusing on proactive monitoring, capacity planning, and rapid incident response. Because Cohesity operates in a highly technical space, you will frequently engage with deep-level infrastructure challenges, ranging from storage performance and network latency to complex cloud-native architecture.
Expect a role that demands both breadth and depth. You will be challenged to understand how hardware, operating systems, and networking protocols interact, and you will be expected to apply that knowledge to optimize production environments. It is a demanding position that offers significant strategic influence, as your work directly shapes the stability of the Cohesity ecosystem.



