Site Reliability Engineer (SRE)-Microsoft Hyper-V & Private Cloud

Sandoz Inc.
Houston, TX, United States
21 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours

Tech stack

Data Recovery Disaster Recovery Failover Clustering Monitoring of Systems Hyper-V Windows Servers Performance Tuning Windows PowerShell Reliability Engineering Prometheus System Center Virtual Machine Management Virtualization Technology
+9 more
Private Cloud Environment Cloud Monitoring System Availability Grafana Software Troubleshooting Infrastructure Automation Frameworks Performance Monitor Veeam Splunk

Job description

We are seeking an experienced Site Reliability Engineer (SRE) with expertise in Microsoft Hyper-V, private cloud infrastructure, and virtualization technologies. This role focuses on maintaining highly available, secure, and scalable infrastructure through automation, proactive monitoring, incident management, and performance optimization. The ideal candidate will collaborate with infrastructure, networking, and application teams to improve platform reliability and operational efficiency. Responsibilities Administer and maintain Microsoft Hyper-V, Storage Spaces Direct, and Failover Clustering environments. Support VDI platforms, private cloud infrastructure, and enterprise virtualization services. Automate infrastructure management using PowerShell and other automation tools. Monitor system health, troubleshoot incidents, and perform root cause analysis to improve reliability. Plan capacity, support disaster recovery processes, and optimize infrastructure performance. Requirements, Job Category: Technical Job Description: As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management - Maintain and monitor prod…

  • 1 month ago +

Requirements

Minimum 8+ years of experience in Site Reliability Engineering or Infrastructure Engineering. Strong experience with Microsoft Hyper-V, Windows Server, SCVMM, Storage Spaces Direct, and Failover Clustering. Hands-on experience with PowerShell scripting, infrastructure automation, and VDI environments. Knowledge of enterprise monitoring, networking, capacity planning, and disaster recovery practices. Strong troubleshooting, analytical, communication, and documentation skills. Position Highlights Employment Type: Contract Work Location: Hybrid/Onsite Infrastructure: Microsoft Hyper-V Private Cloud Virtualization: Hyper-V, SCVMM, VDI Automation: PowerShell Monitoring: Azure Monitor, Grafana, Prometheus, Splunk (preferred) Backup & Recovery: Veeam (preferred) Industry Experience: Banking or Financial Services preferred Candidates with strong Microsoft virtualization, automation, and infrastructure reliability experience are encouraged to apply.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

Videos

See all

Related articles

See all