Reliability Engineer
Role details
Job location
Tech stack
Job description
ATP is building a core team of engineers to develop the autonomous skidded production systems that will redefine how energetics are manufactured. As our Reliability Engineer, you will own uptime and reliability of our automated production systems - analyzing failure data, hardening controls and automation architectures, and driving the maintenance and monitoring strategies that keep autonomous plants running continuously with minimal human intervention. The ideal candidate has previously owned reliability or uptime performance on an automated industrial system end-to-end and thinks in terms of failure modes, MTBF, and preventive maintenance., * Own reliability and uptime performance for ATP's autonomous, continuously operating production systems
- Analyze failure data, downtime logs, and maintenance records to identify recurring failure modes across controls, mechanical, and automation subsystems
- Lead FMEA, root cause analysis, and corrective/preventive action efforts for reliability-impacting failures
- Define and track reliability metrics (MTBF, MTTR, availability) across production systems and drive improvement plans
- Work with Controls and Automation engineers to harden PLC/SCADA architectures, sensor redundancy, and fail-safe behavior for continuous, lights-out operation
- Develop predictive and preventive maintenance strategies, including condition monitoring and sensor-based early failure detection
- Partner with Mechanical, Process, and Controls engineers to design out failure points during new system development
- Support commissioning and ramp-up of new production systems with a focus on long-term reliability, not just initial performance
- Lead post-incident reliability reviews and ensure corrective actions are implemented and verified
- Maintain reliability documentation, failure databases, and audit-ready records
- Contribute to test plans, technical documentation, and engineering reviews
- Travel to ATP sites, customer locations, and vendor facilities as needed
Requirements
- Bachelor's degree in Mechanical Engineering, Electrical Engineering, Robotics, Mechatronics, or a related discipline
- 3+ years of professional experience in reliability engineering, maintenance engineering, or automation/controls with a reliability/uptime focus
- Demonstrated experience owning uptime or reliability performance on an automated industrial system end-to-end
- Working knowledge of PLC/SCADA architectures and how automation design choices affect system reliability
- Experience with reliability methods and metrics (FMEA, root cause analysis, MTBF/MTTR)
- Ability to work hands-on in laboratory, fabrication, and plant environments
- Strong problem-solving skills and ability to troubleshoot hardware and automation issues quickly and rigorously, * 2+ years of relevant industry experience
- Experience with predictive maintenance, condition monitoring, or sensor-based failure detection
- Familiarity with Siemens TIA Portal, Ignition SCADA, or equivalent automation platforms
- Familiarity with safety and reliability standards (ISO 13849, IEC 61508, or similar)
- Experience with continuous process systems or other lights-out/autonomous manufacturing environments
- Experience in defense, aerospace, energetics, or other regulated manufacturing environments
Additional Requirements
- Ability to lift 50 lbs
- U.S. Person and ability to obtain a security clearance
ITAR Requirements
This position requires access to information and technology protected under U.S. export control laws and regulations, including the International Traffic in Arms Regulations (ITAR) and the Export Administration Regulations (EAR). To conform to these regulations, applicants must be a "U.S. Person" as defined by 22 CFR § 120.62, which includes (i) U.S. citizens or nationals, (ii) U.S. lawful permanent residents (green card holders), (iii) refugees under 8 U.S.C. § 1157, or (iv) asylees under 8 U.S.C. § 1158.
Benefits & conditions
Compensation Range: $100K - $180K