> Markdown version of [/jobs/ext/2563409-senior-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2563409-senior-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Reliability Engineer - **Company:** ServiceNow - **Location:** Gaithersburg, MD, United States (Remote available) - **Experience:** Expert - **Salary:** $114,700.0 - $195,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon Elastic Compute Cloud, Architectural Patterns, Bash Shell, Software as a Service, Cloud Computing, Cloud Database, Computer Programming, Linux, DevOps, Distributed Systems, Monitoring of Systems, Python (Programming Language), Reliability Engineering, Prometheus, Runbook, Datadog, Scripting, Cloud Platform System, Okta, Snowflake, Grafana, Multi-Cloud, Amazon Virtual Private Cloud (VPC), Cloudformation, Amazon Relational Database Service, Kubernetes, Information Technology, Terraform, Software Version Control, Servicenow, Golang, Microservices - **Published:** August 29, 2026 - **Apply:** https://www.careerjet.com/job/us72ac9e413bbc92f4521db5ed8000a2e0/eaa ## About the Role * US Citizenship Required - role requires or may require federal security clearance * BS degree in Computer Science or related field (or equivalent practical experience) * 7+ years in Site Reliability Engineering, DevOps, or Infrastructure Engineering * Proven track record leading large-scale, cross-team infrastructure projects from conception to production * Demonstrated ability to work autonomously on ambiguous projects with tight deadlines * Hands-on experience with major compliance frameworks (SOC 1/2, ISO 27001, FedRAMP Moderate/High) including infrastructure automation for control implementation and continuous audit evidence collection Technical Expertise * 5+ years with AWS (VPC, EC2, RDS, EKS, CloudFormation) and cloud automation * Expert-level experience with Kubernetes, Helm, Linux, and Terraform * Strong experience with GitOps model, distributed version control, and CI/CD pipelines * Proficiency with monitoring tools (Prometheus, Grafana, DataDog) * Strong programming/scripting skills (Python, Go, Bash) for automation and reading codes * Deep understanding of distributed systems, microservices, and reliability patterns * Experience with Bazel and CueLang a plus * Experience implementing compliance-as-code patterns and automated compliance scanning/remediation Leadership & Communication * Exceptional ability to articulate complex technical concepts to diverse audiences * Track record of driving technical change across organizational boundaries * Successfully delivered multiple complex projects under tight deadlines * Strong customer service orientation with patience and empathy Work Style * Thrives in ambiguous environments and makes progress without perfect information * Hands-on, "can do" attitude with bias for action * Low ego and high intellectual curiosity * Comfortable working across time zones * Self-motivated with strong ownership mentality * Comfortable managing ambiguity in regulatory environments and compliance requirements What Sets You Apart You don't wait for perfect information-you make calculated decisions and drive progress even when requirements are unclear. You've successfully navigated organizational complexity to deliver critical projects on aggressive timelines while building strong relationships across teams and with customers. You view incidents as learning opportunities and naturally raise the bar for technical excellence wherever you go ## Description We are seeking an exceptional Staff Site Reliability Engineer to lead critical infrastructure initiatives and drive innovation across our organization. You'll architect scalable solutions, navigate complex technical challenges independently, and deliver results under tight deadlines in a fast-paced environment. You'll work cross-functionally alongside builders who have helped shape the success of companies such as Google, Okta, AWS, and Snowflake. We are building the next generation identity security platform for the multi-cloud era - will you join us? You will: Strategic Leadership & Technical Execution * Lead enterprise-wide reliability and infrastructure projects across multiple teams with high autonomy * Navigate ambiguous problem spaces and deliver innovative solutions under tight deadlines * Architect and deploy solutions for Cloud Prem and SaaS customers at scale * Drive technical innovation and establish SRE best practices across the organization * Respond to critical incidents, lead root cause analysis, and implement long-term resolutions * Develop automation solutions to streamline operations and reduce manual workload * Participate in on-call rotation and ensure effective incident handoff and documentation * Lead compliance-focused infrastructure initiatives and partner with Security teams on control implementation across cloud environments Cross-Functional Collaboration & Communication * Partner with Engineering, Product, and Customer Success teams to align reliability goals with business objectives * Communicate complex technical concepts effectively to technical and non-technical audiences, including executives * Influence technical decisions across teams through thought leadership and demonstrated expertise * Build consensus and drive adoption of new tools, processes, and architectural patterns Customer-Facing Technical Leadership * Provide tier 2/3 technical support to enterprise customers for complex troubleshooting * Work directly with customer technical teams to resolve deployment, configuration, and integration challenges * Conduct technical onboarding and provide expert guidance on platform architecture and best practices * Create customer-facing documentation, troubleshooting guides, and run-books * Lead customer calls and technical discussions as a trusted advisor Team Development * Mentor SRE and engineering team members, elevating technical capabilities * Foster a culture of reliability, operational excellence, and continuous improvement, We approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. . To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service. Equal Opportunity Employer ServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements. Accommodations We strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact for assistance. Export Control Regulations For positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities. From Fortune. ©2026 Fortune Media IP Limited. All rights reserved. Used under license. By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Top Characteristics of a Software Engineer](https://www.wearedevelopers.com/magazine/166-top-characteristics-of-a-software-engineer) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How to Answer the Interview Question: “Why Do You Want to Be a Software Engineer?”](https://www.wearedevelopers.com/magazine/392-how-to-answer-the-interview-question-why-do-you-want-to-be-a-software-engineer)