Senior System Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
- Ability to drive both Design for Reliability (DfR) and Reliability, Availability & Maintainability (RAM) processes with consistency and rigor for both current and next generation Cloud & AI hardware and infrastructure solutions.
- Work across functionally across different disciplines (Hardware, Firmware, Architecture, Safety, Serviceability etc.) to influence business and operational decisions.
- Works with key stakeholders to set appropriate reliability and availability targets and allocates the targets to lower-level sub-systems and components.
- Develops appropriate tests at system, sub-system and/or/component as needed to demonstrate the budgeted reliability targets.
- Carries out effective Design Failure Modes & Effects Analysis (DFMEA) to identify and mitigate critical risks via design, operational and diagnostic improvements.
- Utilizes different modeling methods such as Reliability Block Diagram (RBD), Markov, Discreet Event Simulations (DES) and Fault Tree Analysis (FTA) to quantify risks and/or support business decisions.
- Develops Prognostics & Health Management (PHM) models for Remaining Useful Life (RUL) prediction based on telemetry data.
Requirements
- Doctorate Degree in Mechanical Engineering, Materials Engineering, Reliability Engineering, Electrical Engineering, or related field AND 2+ years technical engineering experience OR Master’s Degree in Mechanical Engineering, Materials Engineering, Reliability Engineering, Electrical Engineering, or related field AND 4+ years technical engineering experience OR Bachelor’s Degree in Mechanical Engineering, Materials Engineering, Reliability Engineering, Electrical Engineering, or related field AND 5+ years technical engineering experience OR 12+ years relevant technical engineering experience.
Other requirements
Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include but are not limited to the following specialized security screenings: Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud Background Check upon hire/transfer and every two years thereafter., * M.S. in Electrical or Electronic Engineering, Reliability Engineering, or equivalent discipline
- 5+ years of experience performing Reliability Engineering on Complex Systems
- Background in Applied Reliability Engineering statistics for repairable and non-repairable systems.
- Experienced in using software tools such as Reliasoft, JMP and python scripts for performing different reliability analyses.
- Passionate individual who is able to work collaboratively in a team environment and across internal divisions, industry (OEM, ODM), and with customers
- Prior experience driving reliability efforts for Cloud & AI hardware including infrastructure (Networking, Power and/or Cooling).
- Experienced in Prognostics & Health Management (PHM) techniques
- Certified Reliability Engineer (CRE)
Reliability Engineering IC4 - The typical base pay range for this role across the U.S. is USD $119,800 - $234,700 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $160,200 - $261,000 per year.
About the company
Microsoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft’s expanding Cloud Infrastructure and responsible for powering Microsoft’s “Intelligent Cloud” mission. SCHIE delivers the core infrastructure and foundational technologies for Microsoft’s over 200 online businesses including Bing, MSN, Office 365, Xbox Live, Teams, OneDrive, and the Microsoft Azure platform globally with our server and data center infrastructure, security and compliance, operations, globalization, and manageability solutions. Our focus is on smart growth, high efficiency, and delivering a trusted experience to customers and partners worldwide and we are looking for passionate engineers to help achieve that mission.
As Microsoft’s cloud business continues to grow the ability to deploy new offerings and hardware infrastructure on time, in high volume with high quality and lowest cost is of paramount importance. To achieve this goal, the Hardware, Infrastructure Management, and Fundamentals Engineering (HIFE) team is instrumental in defining and delivering operational measures of success for hardware manufacturing, improving the planning process, quality, delivery, scale and sustainability related to Microsoft cloud hardware. We are looking for engineers with a dedicated passion for customer focused solutions, insight and industry knowledge to envision and implement future technical solutions that will manage and optimize the cloud infrastructure.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Highest Paying Tech Companies for Developers
Data Engineer Salary UK
Top-Paying Tech Jobs (with Salaries)
Software Engineer Salary London