DevOps Engineer
Role details
Job location
Tech stack
Job description
We seek a DevOps Engineer to maintain, scale, and enhance reliable SaaS platforms operating in distributed, large-scale environments. This role emphasizes operational excellence, automation, observability, and performance to ensure service availability and scalability. Responsibilities include managing initiatives from start to finish, leading critical incident responses, and collaborating with engineering teams to implement lasting improvements benefiting customers directly.
About the Team
With eight specialized departments, the engineering team functions as a highly collaborative, diverse powerhouse. Each department mission is to deliver seamless and innovative communication solutions. These range from software development and machine learning to quality assurance teams that work to create and maintain Zoom's user-friendly interfaces and robust infrastructure. The team continues to push the boundaries of communication technology, bringing people together regardless of their physical distance.
Responsibilities
- Operating, scaling, and continuously improving software-as-a-service production platforms across distributed environments.
- Designing and implementing uninterrupted solutions to ensure continuous operation for services with an availability target of 99.999%.
- Developing and maintaining disaster recovery strategies across data centers located in various regions.
- Developing and maintaining automation, tools, and scripts to enhance deployment efficiency while minimizing manual tasks and improving overall operational workflows.
- Implementing and improving monitoring, alerting, and observability ensures proactive identification and prevention of potential issues.
- Owning system performance, availability, and scalability for services that directly engage with customers.
- Analyzing system behavior and performance data to pinpoint bottlenecks and uncover optimization opportunities.
- Leading incident response efforts, performing root cause analysis, and implementing lasting remediation measures to address identified issues effectively.
Requirements
- 6-10 years of experience supporting and operating SaaS production systems in DevOps or SRE roles
- Design systems that are highly available and meet uptime targets as high as 99.999%.
- Demonstrate expertise in designing and implementing disaster recovery strategies across multiple regions.
- Demonstrate expertise managing distributed systems within production environments that directly serve customers.
- Demonstrate expertise in handling incidents, managing on-call operations, and conducting thorough root cause analysis to address and resolve issues effectively.
Benefits & conditions
$98,900.00
Maximum:
$228,700.00
In addition to the base salary and/or OTE listed Zoom has a Total Direct Compensation philosophy that takes into consideration; base salary, bonus and equity value.
Note: Starting pay will be based on a number of factors and commensurate with qualifications & experience.
We also have a location based compensation structure; there may be a different range for candidates in this and other locations