Information Technology Specialist (Devp Ops Engineer)
Role details
Job location
Tech stack
Job description
We believe in social justice and the power of diversity, equity, and inclusion, and are committed to fighting longstanding discrimination, racial injustice, and systemic racism to contribute to a more just and equitable future for our workforce, our customers, and our partners.
Introduction
This position is in the District of Columbia Health Benefit Exchange Authority (DCHBX). The mission of the Health Benefit Exchange Authority is to implement a health care exchange program in the District of Columbia in accordance with the Patient Protection and Affordable Care Act (PPACA) of 2010 and the District of Columbia Health Benefit Exchange Authority Establishment Act of 2011, thereby ensuring access to quality and affordable health care to all District of Columbia residents.
The incumbent is responsible for designing, implementing, and maintaining the agency's software delivery infrastructure, Kubernetes platform, CI/CD pipelines, and operational monitoring systems supporting DC Health Link.
Duties and Responsibilities
Designs, builds, and maintains GitHub Actions CI/CD pipelines for all platform services, covering the full delivery chain: automated testing, multi-stage Docker image build, vulnerability scanning, image push, image tag update, and ArgoCD-triggered rolling deployment to EKS. Manages the production image promotion lifecycle and administers the promotion workflows.
Manages all Kubernetes manifests across deployment environments. Administers ArgoCD for GitOps-based continuous delivery; monitors application sync status, resolves drift, and ensures the cluster state reflects the Git-declared state at all times. Enforces the trunk-based development branching model.
Maintains Docker multi-stage build configurations across all services; manages Docker for multi-platform builds; manages GitHub Container Registry (GHCR) image lifecycle and access controls. Administers the production AWS EKS cluster and all supporting AWS infrastructure, managed via Terraform.
Maintains RabbitMQ administration, AMQP port configuration, queue monitoring, and integration with the Microservices event-driven architecture; ensures the event bus remains available as a platform-wide dependency for all services. Serves as technical lead during production incidents; performs triage, coordinates response across development and operations teams, produces blameless post-incident reviews, and implements monitoring or runbook improvements to prevent recurrence
Maintains disaster recovery runbooks and is capable of executing a full platform rebuild from Terraform foundation through EKS cluster provisioning, Kubernetes core component installation, secret creation, and ArgoCD-driven application deployment. Provides technical guidance and mentorship to development team members on DevOps practices, tooling, and operational awareness. Partners with the software development lead to identify and resolve build, deployment, and environment-related obstacles. Documents and presents DevOps processes and tooling improvements to technical and management audiences. Participates in Agile ceremonies and contributes to sprint planning for infrastructure-related work items tracked in Jira.
Requirements
Specialized Experience - Experience that equipped the applicant with the knowledge, skills, and abilities to perform the duties of the position successfully, and this is typically in or related to the work of the position to be filled. To be creditable, at least one (1) year of specialized experience must have been equivalent to at least the next lower grade level in the normal line of progression for the occupation in the organization.
Benefits & conditions
Work is generally performed in an office setting.
Promotion Potential
No known promotion potential.
Other Significant Facts
Tour of Duty: Monday - Friday 8:15 a.m.- 4:45 p.m.
Pay Plan, Series, Grade: CS-2210-15