> Markdown version of [/jobs/ext/2715137-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2715137-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Diligent Corporation - **Location:** New York, United States - **Experience:** Expert - **Salary:** $131,000.0 - $164,000.0 - **Contract:** Permanent contract - **Skills:** Microsoft Access, Active Directory, Artificial Intelligence, User Authentication, Build Automation, Microsoft Azure, Ubuntu (Operating System), CentOS, Software as a Service, Continuous Integration, Linux, Domain Name System (DNS), VMware ESX Servers, Virtual Private Networks (VPN), Python (Programming Language), Network Security, Linux System Administration, Windows Servers, Performance Tuning, PowerCLI, Windows PowerShell, Red Hat Enterprise Linux, Reliability Engineering, Ansible, Transmission Control Protocol (TCP), VMware VSphere, Load Balancing, Firewalls (Computer Science), Git, Build Management, Pure Storage, Data Management, CIS Benchmarks, Terraform, Software Version Control, Big Ip, Cisco, Jenkins, Vmware - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/staff-site-reliability-engineer-diligent-corporation-7449619 ## About the Role * 10+ years of experience in systems or infrastructure engineering, including operating large-scale enterprise or SaaS datacenter environments. * Deep hands-on expertise with VMware vSphere (ESXi, vCenter, DRS, HA, vMotion, distributed switches) in production. * Strong Linux administration skills (RHEL/CentOS/Ubuntu), including performance tuning, system hardening, and advanced troubleshooting. * Solid experience with Windows Server and Active Directory (Group Policy, DNS, authentication and access integrations). * Proven track record building and maintaining automation using PowerShell/PowerCLI, Ansible, Python, or similar tools, plus familiarity with Git or other version control. * Good understanding of storage (SAN/NAS), TCP/IP networking, DNS, VPNs, firewalls, and production monitoring/alerting. * A collaborative, problem-solving mindset with the ability to lead complex incidents, communicate clearly, and operate in an on-call, high-availability environment. It would be great if you had these too, but we'll support you if you don't * Experience with enterprise storage and compute platforms such as Pure Storage or Cisco UCS. * Familiarity with Terraform, Jenkins, Azure DevOps, or similar tools for infrastructure as code and CI/CD automation. * Exposure to security hardening and compliance frameworks such as CIS benchmarks, NIST, or ISO 27001. ## Description You're a seasoned Site Reliability Engineer who loves owning complex infrastructure, making things run faster, safer, and with less manual effort. In this Staff-level role, you'll design and operate VMware-based private cloud platforms that power mission-critical SaaS products used by customers around the world. You'll work across Linux, Windows Server, networking, storage, and automation frameworks to increase reliability, reduce toil, and modernize a global datacenter environment. You'll have the scope to set technical direction, build automation at scale, and mentor engineers while staying hands-on with VMware vSphere, F5/AVI load balancers, and hybrid Active Directory. Here's a breakdown of what you'll do (not all of it, just the important stuff) * Lead the architecture, deployment, and ongoing optimization of VMware vSphere-based private cloud infrastructure across multiple global datacenters. * Design and build automation using PowerShell/PowerCLI, Ansible, Python, and CI/CD tools to streamline provisioning, configuration, and compliance. * Administer, harden, and troubleshoot Linux (RHEL/CentOS/Ubuntu) and Windows Server environments that host enterprise and SaaS workloads. * Integrate and manage Active Directory for authentication, access control, and service accounts across hybrid on-prem and cloud environments. * Partner with network and security teams to manage firewalls, VPNs, storage, and load balancers (F5 BIG-IP, AVI/NSX Advanced Load Balancer) for highly available services. * Document architectures and runbooks, participate in on-call and change management, and mentor engineers while influencing long-term reliability and automation strategy., We are committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, you may contact us at recruitment@diligent.com. To support a fair and consistent hiring process, we use AI and digital automated tools to assist our recruitment team in organizing candidate data, surfacing relevant applications, coordinating interview scheduling and summarizing interview notes. All final screening, evaluation, and hiring decisions are made by humans on our team. To learn more about how we collect, protect, and process your information, please review our Candidate Privacy Policy. At Diligent, we recognize the value of a diverse global workforce. Visa sponsorship may be available for select positions based on business needs and specific role requirements. Sponsorship decisions are evaluated on a case-by-case basis, and eligibility should not be assumed for all opportunities. Recruitment agencies: Diligent does not accept unsolicited agency resumes. Please do not forward resumes to our jobs alias, Diligent employees or any other organization location. Diligent is not responsible for any fees related to unsolicited resumes ## Related Videos - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Navigating the Corporate Jungle: Life as a Developer in a large Company](https://www.wearedevelopers.com/videos/621-navigating-the-corporate-jungle-life-as-a-developer-in-a-large-company) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Computer Vision from the Edge to the Cloud done easy](https://www.wearedevelopers.com/videos/263-computer-vision-from-the-edge-to-the-cloud-done-easy) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)