> Markdown version of [/jobs/ext/2675261-lead-cloud-architect](https://www.wearedevelopers.com/jobs/ext/2675261-lead-cloud-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Cloud Architect - **Company:** Peraton Inc - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $112,000.0 - $179,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Backup Devices, Bash Shell, Cloud Computing, Cloud Engineering, Noise Reduction, Linux, DevOps, Disaster Recovery, Failover, Python (Programming Language), Windows Servers, OpenShift, Windows PowerShell, Red Hat Enterprise Linux, Reliability Engineering, Cloud Services, Ansible, Software Deployment, Datadog, Google Cloud, Grafana, Technical Debt, Gitlab, Gitlab-ci, Kubernetes, Deployment Automation, Terraform, Splunk, Ansible Tower, Dynatrace, Jenkins - **Published:** September 1, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9131382/lead-cloud-architect ## About the Role * Must be a U.S. Citizen with the ability to obtain and maintain the required Public Trust level clearance. * Bachelor's Degree and 12 years of experience, or Master's and 10 years of experience, or a High School diploma/equivalent and 16 years of experience. * 7+ years hands-on experience in cloud engineering, DevOps, or production systems engineering. * Extensive hands-on experience operating in AWS Commercial and AWS GovCloud, including OpenShift (ROSA) or comparable Kubernetes-based platforms * Deep infrastructure-as-code experience with Terraform and Ansible / Ansible Tower. * Advanced proficiency of GitLab and Jenkins CI/CD platforms, including reliability gating and deployment automation. * Deep experience in Linux and Windows Server administration * Substantial practical experience of implementing and managing enterprise observability tools such as Dynatrace, Datadog, Splunk and Open Telemetry. * Ownership of an SLI/SLO and alerting program, including error budgets, alert rationalization, and noise reduction. * Extensive scripting/automation proficiency in Python, Bash, PowerShell, or Go. * Experience operating in federal or regulated environments (FISMA, FedRAMP, NIST 800-53). Preferred Qualifications: * AWS Solutions Architect, AWS DevOps Engineer, or AWS SysOps certification. * Red Hat Certified Specialist in ROSA, Red Hat Certified Advanced System Administrator in OpenShift * Azure Solutions Architect Expert, Azure DevOps Engineer Expert, GCP Professional Cloud Architect, or GCP Professional Cloud Developer certification * Dynatrace Associate, Datadog Log Management Fundamentals certification * GitLab CI/CD Associate certification, Certified Jenkins Engineer (CJE) * Terraform Authoring and Operations Professional certification ## Description Peraton is seeking a Lead Cloud Architect to join a team responsible for the operational reliability and architecture of production systems running in AWS Commercial and AWS GovCloud environments. Some system components are also deployed in Azure and Google Cloud Platform (GCP). The primary production workload runs on Red Hat OpenShift Service on AWS (ROSA). This is a hands-on technical leadership role that combines cloud architecture, software engineering practices, and systems operations. The Lead Cloud Architect will design and implement scalable, secure, and reliable cloud solutions; define and monitor SLIs and SLOs; support application deployments and migrations across staging and production environments; and create and manage CI/CD pipelines. They will also lead incident response, drive automation to reduce operational toil, establish and promote cloud engineering best practices, and mentor junior Cloud Engineers. This person partners closely with platform engineers, the security team, and application developers to ensure the infrastructure services are reliable, available, and deployed in a way that meets both developer and security requirements., * Own the reliability, availability, performance, and operational health of critical infrastructure services and applications, driving continuous improvement throughout their lifecycle. * Lead complex production incidents and service recovery efforts, providing technical direction during high-impact events and driving thorough root-cause analysis and corrective actions. * Define and continuously improve reliability objectives, including SLIs, SLOs, error budgets, service health metrics, and operational standards for infrastructure services and applications. * Lead observability strategy in partnership with application teams, defining meaningful metrics, logs, traces, dashboards, and alerts and ensuring they are integrated into the organization's observability capabilities. * Lead the operational lifecycle of deployed services, including releases, upgrades, patching, configuration changes, capacity management, maintenance, and technology refreshes. * Identify and drive the resolution of systemic reliability risks, recurring failure modes, capacity constraints, technical debt, and other sources of operational risk. * Design and lead resilience and recovery initiatives, including failure-mode analysis, performance and capacity testing, disaster recovery, backup and failover strategies, and recovery validation. * Develop and advance automation and everything-as-code practices to improve consistency, reliability, deployment, recovery, and overall operational efficiency. * Partner with Platform Engineering and application teams on architecture and design, defining operational and reliability requirements and influencing platform capabilities and application architectures. * Mentor junior Cloud Architects and provide technical leadership by establishing best practices, improving operational processes, and advancing the organization's overall reliability engineering maturity. ## Related Videos - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Seriously gaming your cloud expertise: from cloud tourist to cloud native](https://www.wearedevelopers.com/videos/373-seriously-gaming-your-cloud-expertise-from-cloud-tourist-to-cloud-native) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best Paying Jobs in Technology](https://www.wearedevelopers.com/magazine/256-best-paying-jobs-in-technology) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What’s the Difference between a Junior, Mid, and Senior Developer?](https://www.wearedevelopers.com/magazine/238-what-s-the-difference-between-a-junior-mid-and-senior-developer)