Unlock Full Resume Report

New offer - be the first one to apply!

September 22, 2026

Cloud Hardware Development Engineer (Storage)

Mid

135,000 - 220,000 USD/yr

Cupertino, CA

Quick Facts

  • Role: Cloud Hardware Development Engineer (Storage)
  • Focus: Fleet reliability, failure analysis, sustaining engineering, and feed-forward into next-generation storage server design
  • Travel: Occasional (<10%) regional and international travel

Description

Own fleet reliability and sustaining engineering for deployed storage platforms, driving failure analysis, component lifecycle management, and a feed-forward loop from field performance to next-generation design. Correlate fleet failure data across drive interconnects, storage controllers, and power-loss protection circuits, then drive resolution from symptom to permanent fix. Translate findings into next-generation hardware requirements, updated test content, and improved fleet quality metrics.

Responsibilities

  • Own fleet quality metrics post-launch (annualized failure rates, unsellable server rates, and component-level failure modes).
  • Drive root cause analysis on high-volume failure patterns, including connector/cable-driven failures, storage controller faults, and power-loss protection issues; engage suppliers on corrective actions.
  • Correlate field failure data with manufacturing lot, supplier, and test-escape information to identify systemic trends.
  • Manage platform-specific BOM variants across multiple deployed generations.
  • Translate fleet failure patterns into concrete next-generation design requirements; define and add test content based on observed failure modes.
  • Add failure-mode-driven manufacturing test coverage and contribute to data center vetting automation.
  • Define and execute validation strategies from PCBA bring-up through server and rack integration (power sequencing, signal integrity, thermal characterization, NVMe/SSD subsystem performance).
  • Own hardware debug during EVT/DVT/PVT builds; correlate failures across PCIe, power rails, NVMe links, and storage controller subsystems.
  • Triage hardware issues at ODM facilities and in datacenters; perform root cause analysis and implement corrective actions.
  • Collaborate with storage service teams, datacenter operations, and NPI engineering; drive ODM/JDM corrective action implementation; partner with firmware/software/operations teams on diagnostic tooling and fleet dashboards.

Benefits

Comprehensive benefits including health insurance (medical, dental, vision, prescription; Basic Life & AD&D; options for Supplemental life), EAP, mental health support, Medical Advice Line, Flexible Spending Accounts, adoption and surrogacy reimbursement, 401(k) matching, paid time off, and parental leave. The package may include sign-on payments and RSUs; final compensation depends on experience, qualifications, and location.

Requirements

  • Bachelor's degree in Electrical Engineering, Computer Engineering, or equivalent.
  • 2+ years of hardware design, development, and validation experience for server or compute platforms.
  • Experience in one or more server technologies: thermal/mechanical design, power delivery, high-speed signal integrity, or accelerator subsystems.
  • Experience developing functional specifications, design verification plans, and validation test procedures.

Similar jobs you might like