> Markdown version of [/jobs/ext/2880761-site-reliability-engineer-container-platform](https://www.wearedevelopers.com/jobs/ext/2880761-site-reliability-engineer-container-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - Container Platform - **Company:** Stellent IT LLC - **Location:** Jersey City, NJ, United States - **Experience:** Expert - **Salary:** $96,800.0 - $145,200.0 - **Contract:** Temporary to permanent - **Skills:** Amazon Web Services, Application Lifecycle Management, Cloud Computing, Linux System Administration, OpenShift, Reliability Engineering, Cloud Services, Software Deployment, Google Cloud, Containerization, Kubernetes, Information Technology, Rancher, Drilldown, Azure AKS - **Published:** September 13, 2026 - **Apply:** https://www.careerjet.com/jobad/usa039747b0b03c319010976f163de77cc ## About the Role * Bachelor's or Master's degree in Computer Science or a related technical field. ## Description We are looking for an experienced Site Reliability Engineer (SRE) to support and maintain enterprise container platforms across on-premises and public cloud environments, including OpenShift, Rancher/RKE, Azure AKS, AWS, and Google Cloud. The ideal candidate will have strong hands-on experience with Kubernetes/container platforms, Linux administration, cloud infrastructure, automation, monitoring, security, and incident management. This role will focus on improving platform reliability, troubleshooting complex infrastructure issues, reducing operational toil, and driving automation and operational excellence., * Monitor, troubleshoot, and maintain container platforms including OpenShift, Rancher/RKE, and Azure AKS. * Troubleshoot platform performance, connectivity, availability, and security issues. * Perform deep-dive analysis of systemic and latent reliability issues. * Participate in incident and problem management activities. * Identify, analyze, and remediate infrastructure vulnerabilities and application deployment issues. * Conduct blameless Root Cause Analysis (RCA) and work with engineering and operations teams to implement permanent fixes. * Support application onboarding and provide troubleshooting throughout the application lifecycle. * Identify opportunities to automate repetitive operational tasks and reduce TOIL. * Partner with Risk and Compliance teams to implement controls and remediate vulnerabilities. * Ensure platform resiliency during implementations and proactively identify and resolve resiliency gaps. * Work with Architecture, Engineering, and Product teams as a key stakeholder in cloud service design. * Support highly available, multi-datacenter environments. * Participate in 24x7 on-call support using a follow-the-sun model. ## Related Videos - [This Is Not Your Father's .NET](https://www.wearedevelopers.com/videos/967-this-is-not-your-father-s-net) - [Containers in the cloud - State of the Art in 2022](https://www.wearedevelopers.com/videos/410-containers-in-the-cloud-state-of-the-art-in-2022) - [From Zero to Hero: Launch & Manage Your Cloud Apps with Free OpenShift & Red Hat Developer Hub](https://www.wearedevelopers.com/videos/1023-from-zero-to-hero-launch-manage-your-cloud-apps-with-free-openshift-red-hat-developer-hub) - [The Future of Cloud is Abstraction - Why Kubernetes is not the Endgame for STACKIT ](https://www.wearedevelopers.com/videos/413-the-future-of-cloud-is-abstraction-why-kubernetes-is-not-the-endgame-for-stackit) - [From Code to Motion: Building an Autonomous Hat-Hunting Robot with Kubernetes & ML](https://www.wearedevelopers.com/videos/1609-from-code-to-motion-building-an-autonomous-hat-hunting-robot-with-kubernetes-ml) - [Local Development Techniques with Kubernetes](https://www.wearedevelopers.com/videos/166-local-development-techniques-with-kubernetes) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)