Elastic SRE - Security & Observability

Zachary Piper
Hanscom Air Force Base, MA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$180,000.0 - $200,000.0
Working hours
Regular working hours

Tech stack

Cyber Security DevOps Distributed Systems Elasticsearch Uptime Networking Basics Reliability Engineering Ansible Security Information and Event Management Data Logging System Availability Delivery Pipeline
+4 more
SC Clearance Kubernetes Terraform Splunk

Job description

Zachary Piper Solutions is seeking an experienced Elastic Site Reliability Engineer (SRE) to support a high-visibility federal engagement focused on observability, platform reliability, and security operations across classified environments. This position will support mission-critical Elastic infrastructure deployments at Hanscom AFB, MA The ideal candidate will have deep expertise supporting enterprise Elastic Stack environments, Kubernetes-based deployments, and production SRE operations within secure or regulated infrastructure environments., Operate, maintain, and optimize large-scale Elastic Stack environments supporting logging, search, observability, and telemetry operations. Ensure platform reliability, uptime, scalability, and performance across production mission systems. Manage Kubernetes-based Elastic deployments, including ECK operator environments. Develop and maintain automation for deployment workflows, monitoring, alerting, and incident response processes. Integrate Elastic infrastructure with SIEM and security tooling including Splunk, EDR platforms, and telemetry systems. Troubleshoot complex issues across distributed systems, infrastructure, and application environments. Implement and support observability frameworks including logging, metrics, tracing, and monitoring solutions. Support CI/CD pipelines and infrastructure-as-code initiatives within DevOps environments. Maintain operational runbooks, escalation procedures, and technical documentation. Participate in on-call support rotations and incident response activities.

Requirements

5+ years of experience supporting Site Reliability Engineering, DevOps, or infrastructure operations environments. Strong hands-on experience with Elastic Stack in enterprise production environments. Advanced Kubernetes experience, including ECK operator deployments. Strong Linux/Unix administration and networking fundamentals. Experience supporting observability, telemetry, logging, and monitoring platforms. Experience working within secure, classified, federal, or highly regulated environments. Ability to work onsite at Hanscom AFB (MA). U.S. Citizenship with ability to obtain or maintain a Secret clearance.

Nice-to-Haves Elastic certifications including Elastic Engineer, Security, or Observability. Experience with Terraform, Ansible, and CI/CD pipeline automation. Exposure to SIEM and EDR technologies including Splunk, CrowdStrike, or Trellix. Experience supporting GovCloud, DoD, or federal infrastructure environments. Prior experience supporting distributed logging or telemetry platforms.

Soft Skills Strong incident response and operational troubleshooting mindset. Ability to remain calm and effective during production outages or high-pressure situations. Strong collaboration skills across security, infrastructure, DevOps, and operations teams. Excellent communication skills for escalation and operational coordination environments. Self-sufficient and capable of operating independently within classified environments.

Benefits & conditions

Compensation: $180,000 - $200,000 annually. Long-term federal engagement supporting mission-critical infrastructure initiatives. Opportunity to support advanced observability and security operations within classified environments.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:53 min

Transitioning toward DevSecOps with dynamic scanning and secrets management

Christoph Ruggenthaler · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:22 min

Leveraging unique cultural backgrounds in engineering design

Ixchel Ruiz · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

4:35 min

Defining service-level indicators based on user behavior

Maxim Schepelin Maxim Schepelin · WWC Europe 2026

Videos

See all

Related articles

See all