> Markdown version of [/jobs/ext/2270548-aws-cloud-platform-engineer](https://www.wearedevelopers.com/jobs/ext/2270548-aws-cloud-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AWS Cloud Platform Engineer - **Company:** Trebecon LLC - **Location:** Naperville, IL, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon Cloudfront, Amazon Elastic Compute Cloud, Amazon S3, Automation of Tests, Microsoft Azure, Backup Devices, Cloud Computing, Cloud Engineering, Continuous Delivery, DevOps, Disaster Recovery, Domain Name System (DNS), Github, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Network Security, Linux System Administration, Octopus Deploy, Redis, Release Management, Reliability Engineering, Shell Script, Datadog, Autoscaling, Amazon ElastiCache, Istio, System Availability, Multi-Cloud, Amazon Virtual Private Cloud (VPC), Git, Containerization, Git Flow, Kubernetes, Infrastructure Automation Frameworks, Deployment Automation, Performance Monitor, Azure AKS, Route53, Cloudwatch, Terraform, AWS EKS, Docker, Golang - **Published:** August 27, 2026 - **Apply:** https://www.dice.com/job-detail/4ee8bd04-4e89-426c-b315-8aec7ab3d028 ## About the Role 4+ years in DevOps, SRE, Platform Engineering, or Cloud Engineering. Expert-level Kubernetes administration. Strong experience with EKS and/or AKS. Deep understanding of Argo CD and GitOps principles. Strong GitHub Actions experience. Advanced Git branching and release management knowledge. Strong AWS architecture and operational expertise. Terraform expertise. Containerization experience (Docker/Kubernetes). Linux system administration. Cloud networking and security. Preferred Azure experience. Service Mesh technologies. Multi-cloud deployments. Datadog or similar observability platforms. Python, Go, Shell scripting, or automation development. Platform Engineering experience. ## Description Deploy, manage, and troubleshoot workloads across AWS EKS and Azure AKS environments. Design and maintain Kubernetes clusters throughout their lifecycle. Manage Kubernetes networking, ingress controllers, service meshes, pod communication, and traffic routing. Troubleshoot complex issues involving pods, nodes, networking, DNS, storage, and cluster scalability. Implement high availability and disaster recovery strategies for Kubernetes platforms. GitOps & Continuous Delivery Design and manage enterprise-scale GitOps workflows using Argo CD. Automate deployments and environment promotion strategies. Build and maintain CI/CD pipelines using GitHub Actions. Define and enforce Git branching strategies and release management practices. Improve deployment reliability, rollback capabilities, and release governance. AWS Cloud Engineering Strong hands-on experience with: EKS EC2 Route 53 IAM CloudFront/CDN ALB & NLB VPC & Networking Security Groups & NACLs Secrets Manager ElastiCache (Redis) S3 CloudWatch RDS Autoscaling & High Availability Architectures Responsibilities include: Designing secure and scalable cloud architectures. Implementing disaster recovery and business continuity solutions. Optimizing cloud cost, performance, and reliability. Managing multi-account AWS environments. Infrastructure as Code Build and maintain reusable Terraform modules. Provision cloud infrastructure using Infrastructure as Code best practices. Manage environment consistency and compliance through automation. Troubleshoot Terraform state, drift, and large-scale deployments. Platform Automation & Engineering Excellence Continuously identify manual processes that can be automated. Develop solutions for: Deployment automation Infrastructure automation Log monitoring automation Incident response automation Self-service engineering capabilities Build internal developer tooling to improve engineering productivity. Observability & Monitoring Design monitoring and alerting solutions using platforms such as Datadog. Implement automated alert correlation and incident reduction mechanisms. Create dashboards, SLOs, and observability standards. Drive proactive monitoring practices across production environments. Disaster Recovery & Resiliency Design highly resilient cloud architectures. Lead recovery planning for major production outages. Demonstrate expertise in scenarios such as: Region failures Kubernetes cluster failures Account compromises Infrastructure loss events Define recovery strategies using Infrastructure as Code, backups, replication, and automation. ## Related Videos - [Building a Cloud Platform Where Everything is Just Another Kubernetes Resource](https://www.wearedevelopers.com/videos/100137-building-a-cloud-platform-where-everything-is-just-another-kubernetes-resource) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Next-gen CI/CD with Gitops and Progressive Delivery](https://www.wearedevelopers.com/videos/1603-next-gen-ci-cd-with-gitops-and-progressive-delivery) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)