1. What is a Site Reliability Engineer at DoubleVerify?
As a Site Reliability Engineer (SRE) at DoubleVerify, you are at the heart of the infrastructure that powers global digital media measurement. Your work ensures the reliability, scalability, and performance of platforms that billions of data points rely on daily. You will bridge the gap between development and operations, building the systems that allow DoubleVerify to maintain high availability across complex, hybrid environments including GCP, AWS, and on-premises data centers.
This role is inherently strategic. You aren't just "keeping the lights on"; you are actively reducing operational toil through automation, implementing Infrastructure-as-Code (IaC), and integrating cutting-edge AI-assisted development tools to accelerate problem resolution. You will drive incident response for Sev1/Sev2 situations, lead technical projects from planning to deployment, and foster a culture of proactive observability. For a candidate who thrives on solving large-scale distributed systems challenges while directly influencing product stability, this position offers significant technical impact.
