> Markdown version of [/jobs/ext/205884-site-reliability-engineer-sre-cloud-infrastructure-kubernetes](https://www.wearedevelopers.com/jobs/ext/205884-site-reliability-engineer-sre-cloud-infrastructure-kubernetes). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (SRE) - Cloud Infrastructure & Kubernetes - **Company:** JMS Technical Solutions - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $124,800.0 - $145,600.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon Elastic Compute Cloud, JIRA, Bash Shell, Software as a Service, Cloud Computing, Cloud Engineering, Configuration Management, Databases, Continuous Integration, DevOps, Distributed Systems, Github, Identity and Access Management, Python (Programming Language), Linux System Administration, Reliability Engineering, Site Reliability Engineering Practices, Cloud Services, Ansible, Prometheus, Datadog, Circleci, Data Logging, Scripting, Grafana, Kubernetes Helm Charts, Kubernetes, Infrastructure Automation Frameworks, Terraform, Jenkins - **Published:** May 27, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=6af01cf9d066ee92 ## About the Role Do you have experience in RDS database?, * Passion for reliability, customer success, and operational excellence. * Ability to troubleshoot complex distributed systems and quickly identify root causes. * Strong communication skills-able to clearly convey technical concepts to both peers and customers. * A proactive mindset, looking for opportunities to improve processes and prevent issues before they occur. * Flexibility to adapt to changing priorities and technologies, * 3-5 years of experience applying DevOps or SRE practices to production systems. * 2+ years of experience operating workloads in AWS, with a focus on EKS, EC2, IAM, and networking. * 2+ years working with Kubernetes (preferably in production) and Helm. * Experience with IaC tools such as Terraform and configuration management tools like Ansible. * Familiarity with CI/CD pipelines (GitHub Actions, Jenkins, CircleCI, etc.). * Proficiency in scripting languages such as Python or Bash. * Comfortable working in Linux-based environments. Applicants must be authorized to work in the U.S. ## Description As a Site Reliability Engineer (SRE) on the Nautobot Cloud Engineering team, you will help deliver and maintain our managed Nautobot SaaS offering. Your primary focus will be operating, supporting, and evolving customer environments in AWS-especially EKS, EC2, and related services-while ensuring uptime, performance, and security. You will also handle occasional escalations for legacy customers running on AKS or on-premises deployments. This role combines operational excellence with a mindset for continuous improvement. You will work across infrastructure, CI/CD pipelines, and observability tooling, applying DevOps best practices to deliver a reliable, scalable, and secure platform for our customers. A DAY IN THE LIFE: * Operate and support Nautobot Cloud deployments in AWS, including EKS, EC2, RDS, and associated services. * Use Jira to manage operational and project-related tasks, track incidents, and document changes. * Support resolution of escalated issues related to other Kubernetes-like, including AKS or on-prem, customers as needed. * Deploy and update Nautobot instances using Helm charts, Kubernetes manifests, and automation workflows. * Automate improvements to CI/CD pipelines (GitHub Actions, Terraform, Ansible) for provisioning, upgrades, and configuration management. * Maintain observability tools (Prometheus, Loki, Grafana) to ensure accurate monitoring, alerting, and logging. * Troubleshoot application and infrastructure issues across containerized environments. * Collaborate with engineers across Cloud Operations, Nautobot Core, and Nautobot Apps teams to deliver cross-functional solutions. ## Related Videos - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [Integrate your Cognitive Assistant with 3rd-party DBs and software](https://www.wearedevelopers.com/videos/249-integrate-your-cognitive-assistant-with-3rd-party-dbs-and-software) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)