Server Infrastructure Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+54 more
Job description
- Provide advanced L3 support for complex server hardware issues.
- Perform root cause and forensic analysis.
- Manage enterprise server platforms including HPE, Dell, and Lenovo.
- Lead automation initiatives utilizing PowerShell, Python, Ansible, REST APIs, and Redfish.
- Oversee firmware governance, hardware testing and certification.
- Support server design, virtualization, and hybrid/cloud infrastructure solutions.
- Drive monitoring and capacity planning efforts using Dynatrace and Netcool.
- Support security initiatives including SSL certificate deployment and MFA implementation.
- Contribute to enterprise architecture roadmaps.
- Lead technical projects, mentor team members, and establish infrastructure best practices., Advanced Support & Troubleshooting
- Provide Global L3 hardware support and perform in-depth root cause and forensic analysis for complex server issues.
- Serve as a Subject Matter Expert (SME) for:
-
Out-of-band (OOB) management technologies including: o HPE iLO o HPE OneView o HPE COM o Dell iDRAC o Dell OpenManage Enterprise
-
Enterprise server platforms including: o HPE o Dell o Lenovo
- Independently resolve complex, cross-system technical issues with minimal oversight.
Automation & Firmware Management
- Lead scripting and automation initiatives to streamline server deployment, configuration, and management.
- Design and maintain automation solutions using the Redfish API for:
- Server discovery
- Configuration profile management
- BIOS updates
- Firmware lifecycle management
- Develop and maintain automation using:
- Python
- PowerShell
- Ansible
- REST APIs
- Guide application owners in automating application testing and certification processes.
- Own firmware update methodologies and firmware version management for assigned global server fleet segments.
- Identify opportunities to improve automation coverage and firmware management practices across the organization.
Systems Engineering & Testing
- Design and deliver resilient server solutions.
- Lead Failure Mode and Effects Analysis (FMEA) testing initiatives.
- Manage server and component testing, including:
- Planning
- Execution
- Results analysis
- Lead hardware-related projects such as:
- SSL certificate deployment across the server fleet
- Multi-Factor Authentication (MFA) initiatives
- Identify, troubleshoot, and resolve systemic hardware issues while recommending long-term corrective actions.
Documentation & Standards
- Create and maintain comprehensive:
- OOB management standards documentation
- User documentation
- Operational procedures
- Establish and promote infrastructure best practices.
- Contribute to team-wide technical standards.
- Review and improve existing documentation for accuracy, completeness, and consistency.
Infrastructure Design & Engineering
- Design, implement, and maintain enterprise server environments and hybrid/cloud infrastructure solutions that support business operations.
- Research, design, and implement infrastructure solutions for internal customers, including:
- Containerization
- Kubernetes orchestration
- On-demand scaling solutions
- Design and configure secure virtualized environments aligned with enterprise architecture standards, including:
- Virtual machines
- Virtual networks
-
Distributed storage systems o SAN o HDFS o SSD-based storage
- Contribute technical expertise to enterprise architecture roadmaps and infrastructure blueprints.
Monitoring, Capacity Planning & Optimization
- Design, implement, and refine enterprise monitoring solutions using:
- Dynatrace
- Netcool
- Ensure systems are:
- Properly configured
- Running efficiently
- Secure and observable
- Monitor and analyze resource utilization, including:
- CPU
- Memory
- Storage
- Network traffic
- Utilize Dynatrace, Netcool, and related tools to prevent capacity issues and optimize operational costs.
- Configure and maintain Dynatrace dashboards to:
- Monitor hardware firmware versions
- Detect firmware drift from approved versions
- Proactively identify performance bottlenecks.
- Lead capacity planning efforts based on usage trends, forecasting, and analytics.
Team Development & Leadership
- Mentor and guide team members while fostering a culture of technical excellence.
- Lead technical knowledge-sharing sessions.
- Support continuous learning and professional development across the team.
- Independently lead technical projects and initiatives.
Requirements
Seeking an experienced Server Infrastructure Engineer to join our global IT organization. This role is responsible for the design, implementation, support, and optimization of enterprise server infrastructure across a global environment. The ideal candidate will possess deep expertise in server hardware, out-of-band management technologies, firmware lifecycle management, automation, and infrastructure engineering., * Enterprise server hardware expertise (HPE, Dell, Lenovo)
- Out-of-band management technologies (iLO, iDRAC, OneView, OpenManage)
- Windows and Linux server administration
- PowerShell, Python, Ansible, and infrastructure automation
- REST API and Redfish integration
- Firmware lifecycle management and hardware testing
- TCP/IP networking, DNS, DHCP, and routing
- Virtualization, storage, and containerization technologies
- Dynatrace and Netcool monitoring platforms
- Strong troubleshooting, documentation, and technical leadership skills
Preferred Skills:
- Active Directory, Azure, Google Cloud Platform, and Kubernetes
- SAN and distributed storage technologies
- SSL certificate management and MFA solutions
- ServiceNow, BMC, and JIRA
- Performance tuning and capacity planning
- AI-driven automation and operational efficiencies
Experience Required:
- 10+ years of server infrastructure or systems engineering experience
- Minimum 5 years supporting enterprise-scale environments
- Associate degree or equivalent work experience required; Bachelor’s degree preferred, We are seeking an experienced Server Infrastructure Engineer to join our global IT team. This role requires strong technical expertise in server hardware, out-of-band (OOB) management, firmware operations, automation, infrastructure engineering, and enterprise systems design., * Hardware Experience
- Computer Hardware Expertise
Preferred Skills
- Server Administration
- Ansible
- Server Support
- PowerShell
Required Experience
Experience
- 10+ years of experience in:
- Server Infrastructure
- Systems Engineering
- Related IT disciplines
- Minimum of 5 years supporting enterprise-scale environments.
Education
- Associate Degree or equivalent work experience.
Hardware Expertise
- Deep knowledge of x86 server hardware, including:
- CPU architecture
- Bus architecture
- Expansion cards
- Hardware diagnostics and troubleshooting
Server Management
- Expert-level knowledge of:
- OOB management platforms (iLO, iDRAC/DRAC)
- HPE OneView
- HPE COM
- Dell OpenManage Enterprise
- Enterprise hardware lifecycle management
Operating Systems & Networking
- Extensive experience with:
- Windows Server
- Linux Server
- Strong TCP/IP networking knowledge, including:
- DNS
- DHCP
- Broadcast domains
- Routing protocols
Automation & DevOps
- Advanced scripting and automation experience using:
- PowerShell
- Python
- Ansible
- Comparable technologies
- Experience with:
- CI/CD development
- Automated server provisioning and builds
- REST API integrations
Systems Engineering
- Experience with:
- FMEA methodology
- Firmware lifecycle management
- Hardware testing and certification
Virtualization & Storage
- Working knowledge of:
- Virtualization technologies
- SAN storage
- HDFS
- SSD-based storage solutions
- Containerization
Monitoring & Observability
- Hands-on experience with:
- Dynatrace
- Netcool
- Enterprise monitoring and alerting platforms
Leadership & Professional Skills
- Demonstrated ability to:
- Lead technical projects independently
- Mentor team members
- Collaborate across technical and business teams
- Strong analytical, troubleshooting, documentation, and problem-solving skills.
- Excellent written and verbal communication skills.
Preferred Experience
Architecture & Cloud
- Knowledge of:
- Microsoft Active Directory
- Google Cloud Platform (Google Cloud Platform)
- Microsoft Azure
- Kubernetes
- Cloud-native and hybrid cloud solutions
- Familiarity with:
- Storage systems
- Networking infrastructure
- Data center facilities
- Enterprise systems architecture
- Basic cloud computing expertise.
- Performance tuning and optimization experience.
Protocols & Security
- Knowledge of:
- SNMP
- SMTP
- MFA (Multi-Factor Authentication) methodologies
- Experience with:
- SSL certificate management
- Security best practices
IT Operations & Management
- Experience with ITSM and workflow tools such as:
- BMC
- ServiceNow
- Experience using JIRA for project and workload management.
- Knowledge of leveraging AI for:
- Automation
- Troubleshooting
- Documentation
- Workflow design
- Experience with:
- Internal certification testing methodologies
- Server test lab management
- Server decommissioning processes
- Enterprise software license management
- Vendor management
Education Requirements
Required
- Associate Degree
Preferred
- Bachelor’s Degree, * Knowledge of MS Active Directory, Google Cloud Platform, MS Azure, Kubernetes, and other cloud-like solutions.
- Familiarity with storage systems, networking, data center facilities, and overall systems architecture.
- Basic knowledge of cloud computing.
- Performance tuning expertise.
Protocols & Security
- Knowledge of SNMP, SMTP, and MFA (Multi-Factor Authentication) methodologies.
- Experience with SSL certificate management.
IT Operations & Management
- Experience with IT ticketing and workflow systems (e.g., BMC, ServiceNow).
- Experience using JIRA for workload and project tracking.
- Knowledge of using and leveraging AI for automation, troubleshooting, documentation, and workflow design.
- Internal certification testing methods.
- Server test lab management.
- Experience with server decommissioning and enterprise license management.
- Vendor management experience.
About the company
Everforth Apex is a world-class IT services company that serves thousands of clients across the globe. When you join Everforth Apex, you become part of a team that values innovation, collaboration, and continuous learning. We offer quality career resources, training, certifications, development opportunities, and a comprehensive benefits package. Our commitment to excellence is reflected in many awards, including ClearlyRateds Best of Staffing in Talent Satisfaction in the United States and Great Place to Work in the United Kingdom and Mexico.
Everforth Apex uses a virtual recruiter as part of the application process. Click for more details. By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Everforth Apex and its affiliates, and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy at
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
The Most Popular IT Jobs on the Market
Fully Remote Software Engineer Jobs
What Are The Top Skills Required For Azure Developers?
Highest Paying Tech Companies for Developers