> Markdown version of [/jobs/ext/1416873-cloud-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1416873-cloud-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloud Senior Site Reliability Engineer - **Company:** Hispanic Technology Executive Council - **Location:** Jersey City, NJ, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Microsoft Windows, Active Directory, Amazon Web Services, Microsoft Azure, Cloud Computing, Continuous Integration, Dynamic Host Configuration Protocol, Linux, Integrated Windows Authentication, Domain Name System (DNS), Monitoring of Systems, Identity and Access Management, Python (Programming Language), Kerberos (Protocol), Network Security, Log Analysis, OpenShift, Platform as a Service (PAAS), Reliability Engineering, Cloud Services, Ansible, Prometheus, Shell Script, Software Deployment, HybridCloud, Git, Kubernetes, Information Technology, Azure AKS, Elastic Beanstalk, Terraform, Splunk, Dynatrace, Jenkins - **Published:** July 24, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/74519954/1 ## About the Role * BS /MS degree in Computer Science or related technical field involving systems or equivalent practical experience. * Minimum 8+ years of hands-on experience supporting Kubernetes /Openshift / Container PaaS platform * Experience with Python, Ansible and shell scripting * Kubernetes /Openshift /Terraform certifications are a plus * Strong experience in major services related to Compute, Storage, Network and Security * Experience with monitoring tools like Prometheus and Dynatrace, as well as cloud native tools like Azure Monitor and Log Analytics * Strong understanding and background of working with a complex Active Directory and IAM controls * Advanced knowledge of DNS, DHCP, Kerberos and Windows Authentication * Experience with CI/CD tools git /Jenkins, GitOps model * Excellent understanding of Linux /Windows operating systems administration * Systematic problem-solving approach, sense of ownership and drive * Ability to juggle competing priorities and adapt to changes in project scope. * Excellent interpersonal, organizational and communication (written, verbal, and presentation) skills are a must. * Proven ability to work independently with minimal supervision and as part of a team with direct responsibilities. Desired Job Skills: * Experience in Openshift, managed Kubernetes services such as AKS, EKS, or GKE * Experience in Terraform, ArgoCD, Tekton, and K-native technologies * Experience in agile deployment methodologies (GitOps) * Knowledge of various container runtimes * Familiarity with the operator deployment pattern. * Experience working in a highly available multi-datacenter environment * Experience working with monitoring tools such as Prometheus, Splunk, Dynatrace, Sysdig, or similar tools. * Understanding of cost management, inventory management, FinOps model ## Description We are seeking an experienced Senior Cloud Site Reliability Engineer (SRE) to support and administration of our Hybrid Cloud Container (OpenShift /AKS) platform. Our Cloud Service Reliability Engineers (cSREs) ensure that our Cloud services meet the reliability and uptime requirements of our demanding enterprise customers. This is achieved with, the best engineering practices and resilient design and through a well-defined and effective global on-call rotation that runs 24x7. The role provides opportunity to work with wide range of technologies and unique perspective on how various services (on-prem/off-prem) interact with each other. You will work with colleagues that are as smart, hardworking, and driven as you. You will get an opportunity to work in a team that keeps growing, innovating, and giving you room to be proactive and creative., * Responsible for reliability and support of Container PaaS Platform on-prem/off-prem (Azure /AWS /Google) * Monitor and troubleshoot Container PaaS platform (Openshift) and Azure (AKS) environment performance issues, connectivity issues, security issues, etc. * Perform deep dives into systemic and latent reliability issues, Incident management, problem management * Identifying, analyzing, and resolving infrastructure vulnerabilities and application deployment issues. * Perform blameless RCA, partner with engineering and operation teams across the organization to roll out fixes. * Identify and drive opportunities to improve automation for the PaaS services; scope and create automation for deployment, management, and visibility of our services. * Evaluating and automating the scaling and capacity requirements within PaaS environments * Partner with risk, and compliance teams to bring visibility and implement right controls and policies in the PaaS Platform * Ensure resiliency during implementation and identify/fix resiliency problems by collaborating with engineering teams * Be a key stakeholder in the design of cloud services and work with Architecture, engineering, product teams * Participate in 24x7 on-call coverage follow the sun model ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Kubernetes dev is fun, but setup and ops isn't! See a fun PaaS alternative to push any code, ipynbs or even just data!](https://www.wearedevelopers.com/videos/732-kubernetes-dev-is-fun-but-setup-and-ops-isn-t-see-a-fun-paas-alternative-to-push-any-code-ipynbs-or-even-just-data) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [What’s the Difference between a Junior, Mid, and Senior Developer?](https://www.wearedevelopers.com/magazine/238-what-s-the-difference-between-a-junior-mid-and-senior-developer) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)