Delivery Engineer - Cloud & Automation

Bluestaq LLC
United States
about 1 month ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$95,000.0 - $140,000.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Systems Engineering Bash Shell Software Debugging DevOps Identity and Access Management Python (Programming Language) Linux System Administration Octopus Deploy Windows PowerShell
+13 more
Reliability Engineering Prometheus Datadog Istio Grafana Amazon Virtual Private Cloud (VPC) Kubernetes Low Latency Deployment Automation Apache Kafka Linkerd (Service Mesh) Cloudwatch Terraform

Job description

Bluestaq is hiring a Delivery Engineer to operate and evolve our Kubernetes-based platform on AWS. You will own the deployment automation, day-to-day production health, and incident response for the services that power the Unified Data Library (UDL) - a mission-critical data platform used across government and commercial space operations.

This is a hands-on platform/SRE role. You will spend your time deploying infrastructure with Terraform/Terragrunt, debugging failed rollouts in EKS, chasing latency and error spikes through Grafana and traces, and remediating production issues in Kubernetes directly.

Why This Role Matters

Bluestaq delivers integrated, secure, and reliable systems for customers operating in mission-critical and commercially demanding environments. Systems engineering is at the core of how we win, deliver, and scale. Without strong systems engineering execution, our teams can’t move fast, our systems can’t stay secure, and our customers can’t rely on the technology that supports their operations.

This role is foundational to the team and the organization. Your work spans defined features, environments, and components within a team, with growing independence, which means your decisions, your designs, and your technical judgment have a direct and measurable effect on what Bluestaq delivers. We’re building a culture of ownership and technical excellence, and we need people who bring both depth and the right mindset to do it.

What Success Looks Like in the First 90 Days

  • Comfortable deploying and updating services across our EKS environments via Terragrunt and Argo CD
  • Independently triaging Grafana alerts and resolving common production issues
  • Contributing improvements to runbooks, dashboards, or automation based on at least one real incident
  • Participating in the on-call rotation, * Deploy and maintain UDL platform infrastructure on AWS EKS using Terraform and Terragrunt
  • Manage application rollouts via Argo CD / GitOps workflows; troubleshoot failed deployments, upgrades, and sync issues
  • Monitor production service health using Grafana dashboards, logs, and distributed traces; drive incidents to root cause
  • Perform live remediation in Kubernetes - pod restarts, rolling updates, scaling, node draining, debugging stuck workloads
  • Improve observability coverage, alert quality, and runbooks based on what you learn from real incidents
  • Automate repetitive operational toil using Python, Bash, or PowerShell
  • Partner with product, security, and engineering teams to ship changes safely into regulated environments
  • Mentor junior engineers and share knowledge across the platform team
  • Participate in an on-call rotation for production support
  • You use modern AI tools with judgment: integrating them into workflows where they’re impactful, reviewing their outputs, and handling proprietary or sensitive data responsibly.

Requirements

  • 3-6 years in DevOps, SRE, Platform Engineering, or Infrastructure roles
  • Production experience operating Kubernetes (EKS preferred) - deployments, services, ingress, troubleshooting failed pods/rollouts
  • Hands-on experience with Terraform (Terragrunt a plus) provisioning AWS infrastructure
  • Strong AWS fundamentals (IAM, VPC, EC2, EKS, S3, CloudWatch)
  • Comfort working in Linux environments and writing automation in Python, Bash, or PowerShell
  • Experience with observability tooling - Grafana, Prometheus, OpenTelemetry, or equivalent log/metrics/trace platforms
  • Strong problem-solving skills and clear written/verbal communication, especially during incidents
  • Bachelor’s degree plus 4+ years experience (or equivalent combination of education and experience)

Preferred Experience

  • Argo CD or other GitOps tooling at scale
  • Helm and/or Kustomize for Kubernetes manifest management
  • Kafka / Amazon MSK operations
  • Service mesh experience (Istio, Linkerd)
  • Incident response leadership and postmortem authorship
  • AWS certifications (Solutions Architect, DevOps Engineer, or similar)
  • Experience operating in regulated, security-conscious environments (FedRAMP, IL4/5, DoD)

Additional Requirements

  • U.S. citizenship required (clearance-eligible position)
  • On-site presence in Colorado Springs required
  • May require eligibility for TS/SCI security clearance

About the company

About BluestaqAt Bluestaq, we build secure data platforms that matter for space missions, national defense, healthcare systems, and commercial innovation. Founded in 2018, we’ve become a leader in enterprise software and secure data management by staying focused on what counts: modern architecture, operational excellence, and mission impact.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:32 min

Shifting to a DevOps career from non-technical backgrounds

Megha Kadur · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all