Reliability Engineer, R&D
Own reliability engineering for AI data center reference designs. Build RAM models, run cross-discipline FMEAs, analyze field failures, and drive design changes for high-availability infrastructure.
About the job
Role Scope
Own reliability engineering for the reference design: availability modeled, weak points found, and fixes engineered before deployment. Build the RAM models: failure rates, redundancy, and maintainability quantified per configuration. Run FMEAs across disciplines: failure modes cataloged and designed out with the engineering teams. Close the loop with the fleet: field failures fed back into models and design changes.
What We're Looking For
You've done reliability engineering for infrastructure, energy, or complex hardware. You've built availability models decision-makers used. You've led cross-discipline FMEAs that changed designs. You mine field data for the truth about failure rates. You argue redundancy trade-offs in dollars and nines.
Bonus: Data center topologies. RAM modeling software. Weibull analysis. Maintenance strategy design.
Skills
Reliability Engineering, Availability Modeling, Fmea, Ram Modeling, Weibull Analysis, Failure Rate Analysis, Redundancy Analysis, Field Data Analysis, Maintenance Strategy
Similar jobs
Hardware Engineering jobsLeads RF and cellular hardware integration for autonomous drone products, covering wireless architecture, coexistence, validation, certification, and production. Requires an advanced electrical engineering degree, 7+ years of complex RF product experience, and deep LTE/cellular expertise.
Own geotechnical engineering for 50GW+ data center development: lead subsurface investigations, interpret data for foundation and ground improvement decisions, perform fatal-flaw site screening, and coordinate with structural/civil/environmental teams across multiple states.
Own the architecture and deployment of dense, liquid-cooled rack-scale AI compute products while leading electrical, mechanical, thermal, and integration engineers. The role requires 10+ years of rack-scale systems architecture experience and a record of shipping NVL-class or equivalent products.
Leads architecture, modeling, protection, and safety-case development for MVDC distribution systems powering gigawatt-scale AI campuses. Requires 12+ years in power systems or power electronics, direct MVDC/HVDC design experience, EMT simulation expertise, and technical team leadership.
Leads end-to-end architecture and hardware development for medium-voltage solid-state transformers supporting gigawatt-scale AI data centers. Requires 15+ years in power conversion, megawatt-scale converter experience, semiconductor strategy expertise, and leadership from architecture through working prototypes.