Data Center Engineer I

Brakebush Brothers
Madison, WI, United States
about 1 month ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Bash Shell Ubuntu (Operating System) Configuration Management Databases Cyber Security Data Centers Data Integrity Disaster Recovery Firmware Issue Tracking Systems Networking Hardware Performance Tuning Windows PowerShell
+13 more
Ansible Runbook VMware Infrastructure Virtualization Technology VMware VSphere System Availability Git Infrastructure Automation Frameworks Information Technology Patch Management Terraform Software Version Control Servicenow

Job description

Summary: This position is primarily responsible for the lifecycle, performance, and reliability of server, SAN, and VMware infrastructure across all facilities. Maintains and optimizes virtualization and backup environments to ensure system availability and data integrity. Partners with IT Security to plan, execute, and validate disaster recovery strategies.

Hybrid Work: Following a mandatory 2-week, on-site training period at the Westfield, WI facility, this role will be primarily work-from-home. Candidates must reside within a reasonable commuting distance of Fitchburg, WI to ensure reliable access to the primary data center and headquarters. This position requires occasional travel to other company sites in Wisconsin and other states as project or operational needs dictate.

Essential Functions:

  • Research, evaluate, and recommend server, storage, and hypervisor upgrades aligned with performance, security, and hardware lifecycle standards.
  • Maintain and optimize data center infrastructure across 8 sites, executing consistent firmware updates, patch management, and configuration compliance across all sites.
  • Develop, maintain, and update technical runbooks, configuration documentation, and CMDB and asset records to ensure accurate, accessible infrastructure knowledge.
  • Diagnose and resolve infrastructure incidents via monitoring alerts and ITSM ticketing systems, performing root cause analysis and extended validation per SLA.
  • Design and implement automation scripts (PowerShell, Bash) to streamline provisioning, patching, and operational workflows.
  • Manage enterprise backup operations using Rubrik, including system health checks, scheduled jobs, and disaster recovery drills in coordination with IT Security and other IT departments.
  • Monitor environmental controls (power, cooling, rack density) and coordinate with Facilities to implement changes that maintain optimal data center operating conditions.
  • Partner cross-functionally with IT Security, Networking, Facilities, and leadership to align infrastructure changes with business, security, and compliance requirements.
  • This position participates in a structured on-call rotation (7-day shifts, beginning Wednesday; 1 week on / 3 weeks off) to support critical infrastructure monitoring, emergency break/fix scenarios, and time-sensitive maintenance windows. Candidates should expect occasional after-hours or weekend communications aligned with incident severity and SLA requirements. All on-call duties fall within the scope of this exempt position and do not trigger overtime eligibility. Call handling follows a defined escalation path via ServiceNow, with guaranteed rest periods between rotations to ensure sustainable coverage.

Requirements

  • Bachelors’ Degree in Information Technology or related field and three plus years of Data Center Engineering experience in a manufacturing setting or five plus years of Server Engineering experience in a manufacturing setting.
  • Hands-on operational experience administering virtualization platforms (VMware vSphere) and enterprise backup infrastructure (Rubrik or equivalent), including configuration, patching, and performance tuning.
  • Demonstrated ability to manage infrastructure projects using change control, risk assessment, and ITIL-aligned processes.
  • Strong technical communication skills, oral and written, with the ability to author clear runbooks, incident reports, and configuration documentation, and coordinate effectively with cross-functional IT, security, facilities teams, and end users.
  • Proficiency in PowerShell/Bash scripting for infrastructure automation and administration tasks., * Experience in regulated manufacturing, food production, or high-throughput industrial environments where strict uptime and compliance standards are required.
  • Advanced VMware vSphere administration: HA/DRS, vRealize, or capacity planning.
  • Advanced Rubrik administration: policy design, ransomware protection, and performance tuning.
  • Proficiency in Ubuntu Linux administration.
  • Experience with Infrastructure as Code (IaC) tools such as Ansible or Terraform, and version control (Git).
  • Familiarity with data center modernization/migration projects.
  • ITIL Foundation certification or equivalent IT service management experience.

Supervisory Responsibility: None

Work Environment: This role operates primarily from a remote home office, with required on-site visits to data centers, corporate headquarters, and production facilities in Westfield/Fitchburg, WI, and occasionally out-of-state locations. Work environments will alternate between a quiet, professional remote setting and active data center/industrial facilities featuring elevated noise levels, temperature variations, and restricted access zones. All personnel must comply with site-specific safety protocols, including wearing required personal protective equipment (PPE) when in production or data center areas.

Physical Demands: The physical demands described here are representative of those required to successfully perform the essential functions of this role. While performing this position, the employee will regularly sit for extended periods using computers and peripheral devices and frequently communicate via phone or video conference. The role requires occasional travel to multiple on-site locations. When on location, the employee must regularly lift, carry, push, and/or pull up to 50 lbs. The position frequently requires kneeling, crouching, crawling, and working in confined rack spaces to install, maintain, or remove server, storage, and networking hardware. Safe ladder operation may be required to reach overhead rack positions. Work is frequently performed in data center and industrial production environments, necessitating exposure to elevated noise levels, temperature variations, and adherence to site-specific PPE requirements.

Travel: This position requires up to 15% travel annually to corporate headquarters, data centers, and production facilities in Wisconsin and other states. Travel is primarily driven by hardware deployments, system upgrades, and break/fix scenarios requiring on-site resolution. While travel is generally scheduled with advance notice, unscheduled trips may be required to address critical infrastructure issues or emergency system outages. A valid driver’s license and REAL ID-compliant identification are required for all travel. Candidates must be prepared for occasional multi-day or extended-duration trips when operational demands require it.

Other Duties: Please note this job description is not designed to cover or contain a comprehensive listing of activities, duties or responsibilities that are required of the employee for this job. Duties, responsibilities, and activities may change at any time with or without notice.

Successful completion of a pre-employment drug test and background check are required.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:28 min

Recognizing vital enterprise stability in legacy software development roles

Gunnar Grosch · Coffee With Developers

3:19 min

Executing complex workflows using Ansible Automation Platform

Goetz Rieger Goetz Rieger · World Congress 2025

Videos

See all

Related articles

See all