Reliability Engineer

Vanderhouwen & Associates, Inc.
Albuquerque, NM, United States
15 days ago
Apply on www.vanderhouwen.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Systems Engineering Audit Trail Build Automation Automation of Tests Configuration Management Continuous Integration DevOps Identity and Access Management Key Management Octopus Deploy Software Architecture
+13 more
Reliability Engineering Software Deployment Policy as Code Cloud Platform System Cloudformation Containerization Kubernetes Information Technology Deployment Automation Terraform Docker Security Orchestration, Automation & Response Vulnerability Analysis

Job description

  • Own and maintain the end-to-end CI/CD pipeline, including build automation, automated testing gates, artifact management, release processes, and production deployments.
  • Lead software deployments into secure, access-controlled operational environments, including hands-on work at customer facilities.
  • Serve as a primary technical resource during site deployments, coordinating with infrastructure, security, and technical teams to resolve issues and validate successful implementations.
  • Design, maintain, and improve infrastructure supporting both development and production environments, with an emphasis on reliability, scalability, security, and operational efficiency.
  • Drive infrastructure-as-code and automation practices to create consistent, reproducible, and maintainable environments while reducing manual operational effort.
  • Establish deployment standards, operational procedures, and rollback strategies to ensure releases are documented, repeatable, and recoverable.
  • Implement and maintain containerization, orchestration, monitoring, alerting, and observability capabilities across production environments.
  • Define service reliability objectives and implement the monitoring and operational practices necessary to measure and maintain system performance.
  • Implement security controls for infrastructure and deployment processes, including identity and access management, secrets management, audit logging, and access controls.
  • Automate security, compliance, vulnerability, and operational checks within CI/CD pipelines and deployed environments.
  • Partner with software architecture and engineering teams on infrastructure decisions involving reliability, scalability, cost, security, and deployment strategy.
  • Coordinate technical requirements, deployment schedules, access needs, and other implementation activities with external stakeholders and site personnel.
  • Develop and maintain infrastructure documentation, deployment procedures, operational runbooks, and technical standards to support continuity and knowledge transfer.
  • Contribute to technical planning, estimation, and continuous improvement initiatives for infrastructure and deployment workstreams.

Requirements

Our client is seeking a Senior Reliability Engineer to own CI/CD, infrastructure, deployment, and operational reliability for a mission-critical software platform operating within secure environments. This role is ideal for an experienced DevOps, site reliability, or infrastructure engineer who is equally comfortable building automated cloud infrastructure and coordinating hands-on deployments at customer sites. The ideal candidate will bring strong technical ownership, operational discipline, and communication skills to ensure reliable, secure, and repeatable deployments., * Active Secret or TS/SCI security clearance.

  • Bachelor’s degree or equivalent professional experience in Computer Science, Systems Engineering, Information Technology, or a related technical discipline.
  • 5+ years of experience in DevOps, site reliability engineering, infrastructure engineering, or a related discipline involving production deployments.
  • Demonstrated experience building, maintaining, and operating CI/CD pipelines from source code commit through production deployment.
  • Strong AWS infrastructure experience, including production cloud environments, containerization with Docker, ECS, or EKS, and infrastructure-as-code technologies such as Terraform or CloudFormation.
  • Hands-on experience with production monitoring, observability, alerting, troubleshooting, and reliability practices.
  • Experience deploying software into external, customer, partner, or other environments outside of internally controlled infrastructure.
  • Strong understanding of deployment automation, infrastructure security, access management, configuration management, and operational best practices.
  • Excellent communication and coordination skills with the ability to work effectively with engineering teams, security personnel, customer stakeholders, and program leadership.
  • Ability to travel and work on-site at customer facilities as required to support deployment and operational activities.

Preferred:

  • Experience deploying or supporting technology within classified, air-gapped, SCIF, or similarly restricted environments.
  • Experience supporting Department of Defense, military, federal government, or defense contractor programs.
  • Experience with Kubernetes and GitOps technologies or workflows such as Argo CD or Flux.
  • Familiarity with high-availability and zero-downtime deployment strategies.
  • Experience implementing security automation, including policy-as-code, vulnerability scanning, compliance checks, or audit automation.
  • Familiarity with security and compliance requirements associated with FedRAMP, ITAR, classified systems, or comparable controlled environments.
  • AWS DevOps Engineer - Professional, Solutions Architect - Associate, or comparable cloud certification.
  • Located in or willing to relocate to the Albuquerque, New Mexico area.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.vanderhouwen.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

2:27 min

Structuring agile and interdisciplinary engineering teams

Oliver Zimmert · LIVE

Videos

See all

Related articles

See all