Site Reliability Engineer

VDart, Inc.
Englewood, NJ, United States
17 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$59,150.0 - $106,925.0
Working hours
Shift work

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Bash Shell Software Debugging DevOps Distributed Systems Fiddler (Software) Groovy Monitoring of Systems Apache JMeter Python (Programming Language) Nginx
+16 more
Reliability Engineering Ansible Akamai Systems Architecture Web Platforms Datadog Data Logging Scripting Performance Testing Infrastructure Automation Frameworks Deployment Automation Video Streaming Terraform New Relic (SaaS) Appdynamics Microservices

Job description

  • Support and enhance observability (monitoring, logging, alerting) across production systems
  • Help maintain SLIs/SLOs for key services
  • Participate in evaluating services for production readiness
  • Collaborate with development teams to identify reliability risks and improve system architecture
  • Contribute to automation of operations, including CI/CD pipelines, incident response, and infrastructure provisioning
  • Participate in incident response and on-call rotations for critical services
  • Contribute to post-incident analysis and drive reliability improvements
  • Partner with security, infrastructure, and product teams to support performance, compliance, and operational excellence

Requirements

  • Willingness to work onsite and participate in a 24/7 on-call rotation as needed
  • 5+ years of experience managing and supporting high-traffic digital platforms
  • Strong experience with CI/CD pipelines and deployment automation
  • Experience with cloud platforms such as AWS and/or GCP
  • Solid scripting skills (e.g., Python, Bash, Groovy)
  • Hands-on experience with observability and monitoring tools like Datadog, New Relic, AppDynamics, or similar
  • Understanding of web, mobile, and OTT architectures
  • Experience supporting large scale websites, Mobile and OTT applications, microservices, APIs, and distributed systems
  • Experience with infrastructure-as-code tools such as Ansible, Terraform, or Chef
  • Familiarity with performance testing tools like JMeter or k6
  • Hands on experience with debugging tools like Charles Proxy or Fiddler

Preferred Qualifications

  • Experience working with CDNs (e.g., Akamai) and reverse proxies (e.g., NGINX, Varnish)
  • Exposure to video streaming platforms and Familiarity with application/infrastructure security controls and best practices

Certifications in SRE, DevOps, or Performance Engineering are a plus

Benefits & conditions

  • $59,150-106,925 per year Description Looking for an opportunity to make an impact? At Leidos, we deliver innovative solutions through the efforts of our diverse and talented people who are dedicated to…

  • 1 day ago +

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

4:55 min

Discussing modern Java language evolution and syntax innovations

Daniel Strmečki · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

7:28 min

Constructing a new Docker layer from scratch

Oliver Seitz Oliver Seitz · World Congress 2026 Europe

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

4:17 min

Generating static microsites for technical documentation using DocToolchain

Johannes Dienst · LIVE

Videos

See all

Related articles

See all