> Markdown version of [/jobs/ext/1248208-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1248208-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Red Hat - **Location:** Fort Liberty, NC, United States (Remote available) - **Experience:** Expert - **Salary:** $118,600.0 - $195,680.0 - **Contract:** Permanent contract - **Skills:** Computer-Aided Design, Application Configuration Access Protocols, Systems Engineering, Authentication Protocols, Microsoft Azure, Cloud Computing, Cyber Security, Computer Programming, Continuous Integration, Data Centers, Linux, DevOps, Distributed Systems, Domain Name System (DNS), Hypertext Transfer Protocols (HTTP), HP Systems Insight Manager, Python (Programming Language), Lightweight Directory Access Protocols (LDAP), Linux System Administration, Open Source Technology, OpenShift, Platform as a Service (PAAS), Red Hat Enterprise Linux, Reliability Engineering, Prometheus, TCP/IP, Private Cloud Environment, Datadog, Scripting, Load Balancing, Autoscaling, Grafana, HybridCloud, Gitlab, Kubernetes, Open Network Automation Platform, Splunk, Webhooks, Fedora, Golang - **Published:** July 12, 2026 - **Apply:** https://www.careerjet.com/job/us23004bb2f61adcdf9d2cbf9d607a9466/eaa ## About the Role 5+ years of experience operating production services on Kubernetes / OpenShift 3+ years of programming experience in Python, Go 2+ years of experience of using cloud providers and technologies (Google, Azure, Amazon, etc.) Hands-on experience with Kubernetes / OpenShift, Linux AWS Experience with GitOps workflows for managing infrastructure or application configuration Solid understanding of Linux systems administration (RHEL/Fedora preferred) Understanding of standard networking (TCP/IP, DNS, HTTP/TLS) and authentication protocols (LDAP) Comfort with incident response, on-call responsibilities Ability to work independently with minimal supervision while keeping the team informed Knowledge of SRE principles - SLOs, error budgets, toil measurement The Following Are Considered a Plus Contributions to open-source projects Experience with the Operator SDK or building Kubernetes operators Red Hat will not be providing visa sponsorship for this position. Therefore, in order to be considered for this position, you must have the ability to work without a need for current or future visa sponsorship. ## Description The Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale by enabling customer self-service, making our monitoring system more sustainable, and eliminating toil through automation. In the IT OpenShift team you will have the opportunity to influence the complex challenges of scale which are unique to Red Hat IT managed cloud platform services, while using your skills in coding, operations, and large-scale distributed system design. We develop, deploy, and maintain Red Hat's next-generation mission critical platform across hybrid cloud infrastructures. We are a global team operating on-premise and in the public cloud, using the latest technologies from Red Hat and beyond. Red Hat relies on teamwork and openness for its success. We learn from our failures in a blameless environment to support the continuous improvement of the team. At Red Hat, your individual contributions have more visibility than most large companies, and visibility means career opportunities and growth. Successful applicants must reside in a state where Red Hat is registered to do business. What Will You Do Design, build, and manage our large-scale infrastructure and platform services, including public cloud, private cloud, and datacenter-based Automate cloud infrastructure through use of technologies (e.g. auto scaling, load balancing, etc.), scripting (python and golang), monitoring and alerting solutions (e.g. Splunk, Splunk IM, Prometheus, Grafana, Catchpoint, DataDog etc) Design, develop, and become expert in IT's Red Hat OpenShift offerings by leveraging emerging industry standards Build & support standardized CI/CD platform components using OpenShift Pipelines and Tekton, GitLab to enable multiple application deployments Apply Infrastructure as Code methodologies using GitOps practices with ArgoCD for declarative platform management Breakdown complex engineering efforts into consumable chunks while working with teams to understand deliverables Design and development of software like Kubernetes operators, webhooks, cli-tools Implement and maintain intelligent infrastructure and application monitoring designed to enable application engineering teams Ensure the production environment is operating in accordance with established procedures and best practices Lead escalation support for high severity and critical platform-impacting events Provide feedback around bugs and feature improvements to the various Red Hat Product Engineering teams Design software tests and lead peer reviews to increase the quality of our codebase Help and develop peers' capabilities through knowledge sharing, mentoring, and collaboration Participate in a regular on-call schedule, supporting the operation needs of our tenants Drive sustainable incident response and lead blameless postmortems Work within a small agile team to develop and improve SRE methodologies, support your peers, plan and self-improve, Overview: GovCIO is currently hiring for a Senior Network Automation Engineer to determine and support Kubernetes platform service requirements for Platform as a Service architec… + 8 days ago, Overview: GovCIO is currently hiring for a Senior Cyber Security / DevOps Systems Engineer to engineer, automate, secure, and optimize full-stack enterprise solutions supporting … + 1 month ago ## Related Videos - [From Code to Motion: Building an Autonomous Hat-Hunting Robot with Kubernetes & ML](https://www.wearedevelopers.com/videos/1609-from-code-to-motion-building-an-autonomous-hat-hunting-robot-with-kubernetes-ml) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers)