> Markdown version of [/jobs/ext/1793678-platform-delivery-reliability-engineer-remote-29337](https://www.wearedevelopers.com/jobs/ext/1793678-platform-delivery-reliability-engineer-remote-29337). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Platform Delivery & Reliability Engineer (Remote) - 29337 - **Company:** Enlighten, an HII - Mission Technologies Company - **Location:** Colorado Springs, CO, United States (Remote available) - **Experience:** Expert - **Salary:** $119,574.0 - $195,000.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Amazon Web Services, Application Services, Microsoft Azure, Bash Shell, Big Data, Computer Programming, Data as a Services, Software Debugging, Linux, Distributed Systems, Python (Programming Language), Cisco Nexus Switches, Release Management, Reliability Engineering, Site Reliability Engineering Practices, Software Engineering, Google Cloud, Apache Spark, Gitlab, Gitlab-ci, Kubernetes, Apache Kafka, Apache Nifi, Presto, Terraform, Golang - **Published:** July 14, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9026590/platform-delivery-reliability-engineer-remote-29337 ## About the Role * Clearance Requirement: Must obtain and maintain a U.S. Government Security Clearance, but not required on day one; U.S. Citizenship required. * 9 years relevant experience with Bachelors in related field; 7 years relevant experience with Masters in related field; or High School Diploma or equivalent and 13 years relevant experience. * Deep, hands-on experience deploying, operating, and debugging production Kubernetes clusters and their ecosystem (networking, storage, service meshes, observability, volume management). * Proven record of delivering complex distributed systems into production across many environments or sites: not just building platforms, but landing them with customers. * Elite troubleshooting and analytical skills across the full stack: Linux systems, hosts, networks, security, containers, and application services. * Experience with infrastructure as code (e.g., Terraform) and modern cloud environments (e.g., AWS, Azure, GCP). * Experience with CI/CD pipelines (e.g., GitLab CI) and proficiency in scripting or programming (e.g., Go, Python, Bash). * Working knowledge of SRE practices: monitoring and alerting, incident management, blameless postmortems, and runbook development. * Demonstrated experience creating or improving engineering and delivery processes that other teams actually adopted; you can point to a workflow that exists because you built it. * Excellent verbal and written communication skills; able to translate deep technical issues for engineers, leadership, and customers, and comfortable coordinating a large number of people across organizational boundaries. * Work Location: *Remote or Hybrid. This role is fully remote unless you are located near one of our offices in Columbia, MD; San Antonio, TX; Boise, ID; Greenville, SC; or Augusta, GA, where a hybrid schedule applies. Note: Work models are subject to change based on business needs. * Experience deploying or operating large-scale data platforms and lakehouse technologies (e.g., Spark, Trino/Presto, Kafka, NiFi, object storage, Iceberg/Delta/Hudi) * Experience delivering into DoD, IC, or other federal environments, including STIG-hardened, disconnected, or air-gapped deployments and familiarity with the RMF/ATO process * Experience with Kubernetes Operators/Controllers development * Prior release management, delivery lead, field engineering, or deployment engineering experience on a multi-team program * Understanding of agile software development methodologies and use of standard software development tool suites (e.g., YouTrack, GitLab, Nexus) * DoD 8140 / 8570 compliance certifications may be required in this position as directed by the customer ## Description Enlighten is looking for a Senior Platform Delivery & Reliability Engineer to own the rollout of our data lakehouse platform across a large, multi-site government enterprise, currently ~50 production Kubernetes clusters and growing. This role is a rare hybrid of platform engineer, SRE, and delivery lead. You can deploy the platform, debug anything you encounter in the field, feed what you learn back to the engineering teams, and help fix underlying issues in the code base. Just as importantly, you can step back from any individual issue and fix the system that produced it by building the processes, tooling, and communication channels inside Enlighten that make every deployment faster and less painful than the one before it. You will be a full member of the Infrastructure team, working daily with our Ingest, Query, Application and Testing teams as well as government customers and site personnel. Success in this role looks like: rollouts across the enterprise happen predictably and efficiently, issues found in the field are cataloged, communicated, and resolved quickly, and the friction that slows deployments steadily disappears. #LI-DS1 #Senior Level * Plan, coordinate, and execute deployments and upgrades of the data lakehouse platform across 50 production Kubernetes clusters in customer environments. * Debug and troubleshoot critical issues anywhere in the stack (infrastructure, Kubernetes, platform services, data services, and applications) and drive them to root cause. * Contribute patches, configuration changes, and automation improvements directly back to the platform. * Catalog and triage issues discovered in the field, communicate them clearly to the Infrastructure, Ingest, Query, and Application teams, and maintain a living knowledge base of failure modes, fixes, and runbooks. * Enable and improve site reliability for fielded environments: monitoring, alerting, incident response, and continuous reliability improvement. * Identify friction and dysfunction in how deployments happen, such as unclear handoffs, communication gaps, and repeated manual work; then design, implement, and institutionalize the processes that eliminate them (release checklists, readiness reviews, escalation paths, cross-team communication cadences). * Continuously improve deployment tooling and automation so rollouts become faster, safer, and more repeatable. * Coordinate with a large set of stakeholders (the engineering teams, government programs, security, and site personnel) and keep them informed. * Mentor engineers, both junior and senior, on debugging, deployment, and operational excellence. * Other duties as assigned. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Now is the time for industrialized software development](https://www.wearedevelopers.com/magazine/601-now-is-the-time-for-industrialized-software-development)