> Markdown version of [/jobs/ext/1719855-site-reliability-engineer-ii](https://www.wearedevelopers.com/jobs/ext/1719855-site-reliability-engineer-ii). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer II - **Company:** Akamai Technologies - **Location:** Denver, CO, United States (Remote available) - **Experience:** Experienced - **Salary:** $95,000.0 - $171,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Continuous Integration, Linux, Distributed Systems, Monitoring of Systems, Python (Programming Language), Linux System Administration, Reliability Engineering, Site Reliability Engineering Practices, Prometheus, Akamai, AI Infrastructure, Saltstack, Grafana, Reliability of Systems, Kubernetes, Infrastructure Automation Frameworks, Hardware Infrastructure, Terraform, Serverless Computing - **Published:** July 1, 2026 - **Apply:** https://dejobs.org/x/x/E63DE1A74EB044548BCBFE8F17A153A8/job/ ## About the Role * Have 2+ years of experience in Site Reliability Engineering and a Bachelor's Degree or its equivalent experience * Demonstrate coding ability in at least one programming language (Python or Go) with experience writing automation * Have experience with Linux systems administration and the ability to troubleshoot complex infrastructure issues * Show familiarity with Kubernetes and containerization concepts * Have experience with monitoring and observability tools such as Prometheus, Grafana, or similar * Have exposure to CI/CD pipelines and infrastructure-as-code tools (Terraform, SaltStack, or equivalent) * Show a willingness to learn and grow, with genuine curiosity about AI infrastructure and distributed systems Work in a way that works for you ## Description In this role, responsibilities will include automation, monitoring, incident response, and working collaboratively with skilled team members. Candidates should possess expertise in Linux systems, automation, and SRE practices. Daily activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: * Building and maintaining dashboards, alerts, and monitoring for inference workloads using Akamai's existing observability platform * Writing automation and tooling in Python or Go to reduce operational toil and improve system reliability * Building and improving runbooks for inference-specific operational procedures, integrating into Akamai's existing incident management processes * Contributing to SLO tracking and reporting, identifying trends and areas for improvement * Supporting CI/CD pipeline maintenance, deployment safety checks, and rollback procedures * Collaborating with product engineering teams to troubleshoot complex problems across the stack * Participating in on-call rotations, responding to production incidents, and conducting blameless post-mortems ## Related Videos - [What the Heck is Edge Computing Anyway?](https://www.wearedevelopers.com/videos/593-what-the-heck-is-edge-computing-anyway) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)