Senior Site Reliability Engineer - OpenShift-Based Platform
Role details
Job location
Tech stack
Job description
Red Hat is looking for a Platform Engineer to join its Platform Engineering team! In this role, you will help architect, implement, improve, and support the OpenShift-based platform that runs many of Red Hat's most important multi-tenant Software-as-a-Service (SaaS) and Managed-service offerings. Using your expertise in SRE principles, you will help create an environment where reliability, scalability, and security come first, and are not treated as an afterthought. In this role, you will spend a portion of your time working across teams to define and iterate upon processes for onboarding new managed services at Red Hat and demonstrate good judgment in employing onboarding methods and techniques that can be repeated and iterated upon. You will also contribute to the codebase of command-and-control software that automates the building, deployment, monitoring, and alerting of Red Hat managed services. The remainder will be spent on various other tasks, such as diagnosing issues, planning, documenting, and mentoring. What you will do
- Design, write, and maintain software (primarily in Python and Golang) that automates the deployment, monitoring, and maintenance of Red Hat managed services.
- Onboarding of new services onto our OpenShift-based platform: adhering to cloud-native design principles & best practices to ensure reliability, scalability, and security; contribute to documents, like standard operating procedures (SOPs) and playbooks, that assist in issue resolution and new-service onboarding.
- Proactively utilize AI-assisted development tools (e.g., GitHub Copilot, Cursor, Claude Code) for code generation, auto-completion, and intelligent suggestions to accelerate development cycles and enhance code quality.
- Participate in an Agile Scrum team that scopes, prioritizes, and allocates work items.
- Participate in an on-call rotation that is responsible for responding to service incidents., Red Hat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law. Red Hat does not seek or accept unsolicited resumes or CVs from recruitment agencies. We are not responsible for, and will not pay, any fees, commissions, or any other payment related to unsolicited resumes or CVs except as required in a written contract between Red Hat and the recruitment agency or party requesting payment of a fee. Red Hat supports individuals with disabilities and provides reasonable accommodations to job applicants. If you need assistance completing our online job application, email application-assistance@redhat.com . General inquiries, such as those regarding the status of a job application, will not receive a reply.
Requirements
Do you have experience in Scrum?, * Background writing object-oriented automation software in Python, experience with Golang is only plus
- Background administering production cloud-native services, preferably containerized and deployed via a container-orchestration system like Kubernetes or OpenShift
- Experience diagnosing service failures and carrying out incident response procedures
- Familiarity with Linux operating system and its configuration
- Ability to effectively work in a globally distributed team
- Understanding of computer networking and protocols, including TCP/IP and DNS
- Understanding of computer security and cryptography basics, including certificates, TLS, and credential-storage systems like Vault is a plus
- Familiarity with CI/CD pipeline concepts and systems, like Jenkins and Tekton/Argo is a plus
- Familiarity with observability tools like Prometheus and Grafana, and how to define metrics that can be used to measure service health and reliability is a plus