> Markdown version of [/jobs/ext/2881377-cloud-infrastructure-manager](https://www.wearedevelopers.com/jobs/ext/2881377-cloud-infrastructure-manager). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloud Infrastructure Manager - **Company:** Amazon.com, Inc. - **Location:** Seattle, WA, United States - **Experience:** Experienced - **Salary:** $228,800.0 - $291,200.0 - **Contract:** Temporary to permanent - **Skills:** Amazon Web Services, Amazon Elastic Compute Cloud, Application Release Automation, Build Automation, Cloud Computing, Configuration Management, Computer Programming, Continuous Integration, Software Debugging, Linux, Domain Name System (DNS), Identity and Access Management, Python (Programming Language), Software Engineering, Management of Software Versions, Istio, Delivery Pipeline, Amazon Virtual Private Cloud (VPC), Cloudformation, Containerization, Kubernetes, Functional Programming, Api Design, Cloudwatch, Api Gateway, Terraform, Golang, Microservices - **Published:** September 13, 2026 - **Apply:** https://www.careerjet.com/job/usd38f1be2e449474bb2cb154e5843135e/eaa ## About the Role * Eight or more years in cloud infrastructure, platform, or SRE engineering, including three or more years managing engineers directly as a hands-on manager who still contributes code and troubleshoots production. * Software development proficiency in Go or Python, strong enough to review the team's platform code and contribute to it directly. * Deep AWS experience at production scale, including multi-account structure, IAM, VPC and networking, and core services such as EC2, Lambda, API Gateway, and CloudWatch. * Hands-on infrastructure-as-code background with Terraform, CloudFormation, or both, including module design and state management. * Production Kubernetes experience, ideally EKS, covering cluster operations, workload scheduling, and containerization on microservices infrastructure. * Experience building internal platform abstractions that other engineering teams consume, Linux systems depth including image build pipelines and configuration management, and experience leading a distributed team. * Experience applying AI tooling across the development lifecycle and agentic approaches to infrastructure management. ## Description Team Leadership * Lead an established team of six to eight engineers through a growth phase expected to roughly double headcount, owning hiring, onboarding, and ramp. * Set direction for a networking sub-group covering VPC design, cloud connectivity, service mesh ingress, and DNS. * Partner with the SRE, Developer Experience, and Data teams and the four product teams the platform serves, and own the team's rotating on-call schedule and the operational standard behind it. Software Development and Engineering * Write and ship platform software in Go or Python, along with infrastructure code in Terraform and CloudFormation. * Apply AI tooling across every stage of software development and use agentic solutions to improve how infrastructure is managed and operated. * Review pull requests, debug failing pipelines, and pick up delivery work when the team is stretched or an incident calls for it. Hands-On Engineering * Review pull requests, debug failing pipelines, and pick up delivery work when the team is stretched or an incident calls for it. * Write and ship infrastructure and platform code in Terraform, CloudFormation, and Go or Python tooling. * Cloud Infrastructure and Infrastructure as Code * Own AWS estate management across accounts, covering compute, storage, networking, identity, and the guardrails that keep the environment consistent. * Drive the infrastructure-as-code practice across a mixed Terraform and CloudFormation footprint, including where each tool belongs and how the two converge over time, and hold the standard for change safety, review, and rollback. * Own the Kubernetes (EKS) platform and the shared components running on it, covering scaling, upgrades, service mesh, and workload isolation. Platform Abstractions * Build the provisioning abstractions that let four product teams request and operate infrastructure without holding the full cognitive load of AWS. * Treat the platform as a product whose users are engineers, covering API design, documentation, versioning, and a deprecation path. * Extend paved paths so observability, security, and compliance requirements are satisfied by default, prioritized against measured friction. Infrastructure Operations and FinOps * Own base image strategy, build automation, CI/CD tooling, and release automation for infrastructure and platform components, keeping patching, hardening, and compliance posture current in a regulated financial services environment. * Establish cost visibility, allocation, and accountability across the AWS estate, starting with tagging discipline, rightsizing, and commitment coverage. ## Related Videos - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Program your infrastructure with CDK and TypeScript](https://www.wearedevelopers.com/videos/144-program-your-infrastructure-with-cdk-and-typescript) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Scoring 2000 Products per Request: Performance Pitfalls in Golang](https://www.wearedevelopers.com/videos/2073-scoring-2000-products-per-request-performance-pitfalls-in-golang) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)