Site Reliability Engineer
VDart, Inc.
Englewood, NJ, United States
17 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.careerjet.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$59,150.0 - $106,925.0
Working hours
Shift work
Job source
Tech stack
Application Programming Interfaces (APIs)
Amazon Web Services
Bash Shell
Software Debugging
DevOps
Distributed Systems
Fiddler (Software)
Groovy
Monitoring of Systems
Apache JMeter
Python (Programming Language)
Nginx
+16 more
Reliability Engineering
Ansible
Akamai
Systems Architecture
Web Platforms
Datadog
Data Logging
Scripting
Performance Testing
Infrastructure Automation Frameworks
Deployment Automation
Video Streaming
Terraform
New Relic (SaaS)
Appdynamics
Microservices
Job description
- Support and enhance observability (monitoring, logging, alerting) across production systems
- Help maintain SLIs/SLOs for key services
- Participate in evaluating services for production readiness
- Collaborate with development teams to identify reliability risks and improve system architecture
- Contribute to automation of operations, including CI/CD pipelines, incident response, and infrastructure provisioning
- Participate in incident response and on-call rotations for critical services
- Contribute to post-incident analysis and drive reliability improvements
- Partner with security, infrastructure, and product teams to support performance, compliance, and operational excellence
Requirements
- Willingness to work onsite and participate in a 24/7 on-call rotation as needed
- 5+ years of experience managing and supporting high-traffic digital platforms
- Strong experience with CI/CD pipelines and deployment automation
- Experience with cloud platforms such as AWS and/or GCP
- Solid scripting skills (e.g., Python, Bash, Groovy)
- Hands-on experience with observability and monitoring tools like Datadog, New Relic, AppDynamics, or similar
- Understanding of web, mobile, and OTT architectures
- Experience supporting large scale websites, Mobile and OTT applications, microservices, APIs, and distributed systems
- Experience with infrastructure-as-code tools such as Ansible, Terraform, or Chef
- Familiarity with performance testing tools like JMeter or k6
- Hands on experience with debugging tools like Charles Proxy or Fiddler
Preferred Qualifications
- Experience working with CDNs (e.g., Akamai) and reverse proxies (e.g., NGINX, Varnish)
- Exposure to video streaming platforms and Familiarity with application/infrastructure security controls and best practices
Certifications in SRE, DevOps, or Performance Engineering are a plus
Benefits & conditions
-
$59,150-106,925 per year Description Looking for an opportunity to make an impact? At Leidos, we deliver innovative solutions through the efforts of our diverse and talented people who are dedicated to…
-
1 day ago +
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.careerjet.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
LM
Luis Minvielle
Is Software Engineering Over-Saturated?
over 2 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
21 days ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
over 2 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
EM
Eli McGarvie
Find a Developer Job: 12 Best Job Sites For Developers
over 3 years ago