Manager, Cloud Hardware Development, Cloud compute/gpu/storage server team
Application deadline: Jul 28, 2026. We have two distinct Cloud Hardware and System Development Manager positions open — one leading the storage server team and one leading the AI/ML (GPU-based) accelerator server team. During the interview process, we will assess fit for both positions and align candidates to the team where their experience and interests are the strongest match.
About The Team
This organization is responsible for designing, building, testing, launching, and maintaining a fleet of AI/ML (GPU-based) servers and storage servers for Amazon's web services. Our engineers work with leading-edge technologies, solve challenging problems, influence the industry's roadmaps, and develop unique solutions that are ahead of the pack. We work in an environment that fosters innovation and creativity — we encourage and invest in new directions and ideas that serve our customers better.
The organization comprises Hardware Design Engineers, Systems Development Engineers, and Technical Program Managers, all with the common goal of delivering the best storage and accelerator server fleet possible to our customers. We are located in Seattle and Cupertino, and we work with ODMs and Design Partners globally. We own the full lifecycle of our server platforms: design, build, test, deploy to the data center, launch, and fleet health beyond launch. There is no hand-off — we are accountable from the first architecture decision through every day the server runs in production.
Responsibilities
- Vision & Architecture
- Set the technical vision and multi-generational roadmap for storage or accelerator (AI/ML/GPU-based) server platforms.
- Make architectural bets that differentiate AWS — anticipating customer needs and industry shifts before they become obvious.
- Manage a team of hardware architects in defining server platform architectures that optimize for performance, reliability, cost, and speed of customer adoption.
- Translate deep understanding of customer workloads (storage, AI/ML training, inference) into hardware design decisions.
- Influence the broader AWS hardware strategy through data, conviction, and results.
- Design, Build & Test
- Own server platform development from architecture through detailed design, prototype, build, and qualification.
- Manage a team of engineers responsible for design, build, and launch of systems.
- Lead ODM/JDM and design partner relationships, ensuring our requirements for performance, quality, testability, and diagnostics are met.
- Drive design verification, system validation, and qualification — ensuring platforms meet reliability, performance, and cost targets before deployment.
- Ensure systems are designed for operational excellence from day one — testability, diagnosability, and serviceability are built in, not bolted on.
- Deploy, Launch & Fleet Health
- Own deployment to the data center, launch readiness, and successful ramp into production.
- Drive qualification and readiness milestones, removing technical and organizational blockers to get servers into the fleet.
- Own fleet health beyond launch — your responsibility never ends. Monitor quality, reliability, and customer experience for the life of the platform.
- Drive toward zero-touch operations — building automation infrastructure that detects, diagnoses, and remediates faults before customer impact.
- Build predictive failure detection capabilities using telemetry, error trending, and log correlation.
- Establish and track fleet health metrics (failure rates, MTTD, MTTR, first-time fix rate, predictive accuracy).
- Close the loop between field failures and design improvements in next-generation platforms.
- Team Leadership & Development
- Manage and grow a diverse team spanning hardware engineering, systems development, and technical program management.
- Hire, develop, and retain top talent across multiple engineering disciplines.
- Create an environment where engineers with fundamentally different expertise (hardware, firmware, software, program management) collaborate effectively and challenge each other.
- Set clear goals, remove obstacles, and hold the team to high standards on delivery and quality.
- Coach and develop senior technical leaders — help architects think bigger and help execution-focused engineers see the strategic picture.
- Cross-Organization Collaboration
- Partner with AWS service teams to ensure server platforms meet data path and control path requirements and drive fast adoption.
- Work with supply chain, manufacturing, and datacenter operations teams to deliver at scale.
- Influence peer teams and senior leadership on technical direction, investment priorities, and trade-offs.
- Represent your team's work and roadmap to VP-level and above.
Requirements
- 7+ years of relevant hands-on systems engineering and administrative work in networking, storage systems, or operating systems experience.
- Experience in server development, e.g., compute, AI/ML, storage, edge servers.
- Hands-on experience in designing, developing, and operationally supporting high-volume enterprise servers.
Qualifications
- Experience working with CM/OEM/ODM vendors for design development and manufacturing.
Pay
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location.
- USA, CA, Cupertino: $201,300.00 - $272,400.00 USD annually
- USA, CO, Denver: $175,100.00 - $236,900.00 USD annually
- USA, WA, Seattle: $175,100.00 - $236,900.00 USD annually
Benefits
Amazon offers comprehensive benefits including:
- Health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance, and option for Supplemental life plans).
- Employee Assistance Program (EAP) and Mental Health Support.
- Medical Advice Line.
- Flexible Spending Accounts.
- Adoption and Surrogacy Reimbursement coverage.
- 401(k) matching.
- Paid time off and parental leave.
Learn more about our benefits at https://amazon.jobs/en/benefits.