> Markdown version of [/jobs/ext/1890401-devops-engineer-aws](https://www.wearedevelopers.com/jobs/ext/1890401-devops-engineer-aws). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps Engineer - AWS - **Company:** TensorWave Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Cloud Computing, Continuous Integration, Linux, DevOps, Digital Architecture, Identity and Access Management, Reliability Engineering, Prometheus, Runbook, TypeScript, Data Logging, Load Balancing, High Performance Computing, Autoscaling, System Availability, Grafana, Backend, Git Flow, Kubernetes, Route53, Cloudwatch, Terraform - **Published:** July 31, 2026 - **Apply:** https://jobs.ashbyhq.com/tensorwave/b2aeb755-b488-4fc0-9734-6283e964ec31 ## About the Role * 5+ years in cloud infrastructure, DevOps, SRE, or platform operations * Hands-on AWS experience: VPCs, EC2, S3, IAM, CloudWatch, Route 53, load balancers, security groups, private networking * Proficiency with IaC tooling (Terraform strongly preferred) * Strong Linux fundamentals - networking, process management, storage, troubleshooting * Experience with CI/CD, Git-based workflows, and monitoring/alerting platforms * Clear communicator who can document infrastructure and collaborate across engineering teams, * Experience with AI/ML, GPU, or HPC workloads * Kubernetes on AWS (EKS or self-managed) * Observability platforms: Prometheus, Grafana, Loki, OpenTelemetry, Datadog * AWS cost optimization: right-sizing, savings plans, lifecycle policies, tagging * Startup or high-growth infrastructure environment background ## Description Reposted 12 Hours Ago Remote Hiring Remotely in USA Senior level Remote Hiring Remotely in USA Senior level Design, provision, and operate AWS infrastructure for AI/HPC workloads across environments. Build and maintain IaC, CI/CD, observability, and scaling patterns. Troubleshoot networking, compute, storage, participate in incident response, document architecture and runbooks, and optimize cost and reliability. The summary above was generated by AI, We are hiring an AWS Cloud Engineer to design, provision, optimize, and support the AWS infrastructure powering our AMD GPU AI/HPC platform. This is a hands-on execution role - you'll work closely with Rust backend engineers, TypeScript developers, SREs, and platform teams to keep cloud infrastructure reliable, cost-efficient, and scalable. The goal is simple: reduce cloud bottlenecks and give our engineering teams a solid foundation to build on. What You'll Do * Own the full lifecycle of AWS infrastructure across dev, staging, production, and customer-facing environments - provisioning, scaling, monitoring, security, cost optimization, and decommissioning * Build and maintain Infrastructure-as-Code (Terraform, Pulumi, AWS CDK, CloudFormation) * Implement cloud patterns for high availability, auto-scaling, secure service communication, and customer environment provisioning * Build and maintain CI/CD workflows for cloud infrastructure and hosted services * Improve observability through metrics, logging, alerting, dashboards, and runbooks * Troubleshoot AWS networking, compute, storage, IAM, and deployment issues * Participate in incident response, post-incident reviews, and root cause analysis * Document architecture, operational processes, and best practices ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)