×
Register Here to Apply for Jobs or Post Jobs. X

Sr Technical Program Manager - Hardware, AWS Generative AI & ML Servers

Job in Northern, Floyd County, Kentucky, USA
Listing for: Amazon Inc.
Full Time position
Listed on 2026-09-30
Job specializations:
  • Engineering
    Systems Engineer
Salary/Wage Range or Industry Benchmark: 148700 - 231400 USD Yearly USD 148700.00 231400.00 YEAR
Job Description & How to Apply Below
Sr Technical Program Manager - Hardware, AWS Generative AI & ML Servers

Job :  | Amazon Development Center U.S., Inc.

Amazon operates the world's largest fleet of GPU-accelerated servers powering AI/ML workloads at cloud scale. Our team designs, builds, and operates this fleet — solving systemic hardware issues and building systems that detect and prevent recurrence so customers experience the highest quality of service.

We are seeking a Senior Technical Program Manager to drive end-to-end delivery of GPU-accelerated servers across our global fleet. You will coordinate cross-functional engineering teams spanning hardware, firmware, and software, manage ODM partnerships across multiple continents, and establish closed-loop quality systems that drive continuous improvements. This role requires technical depth to translate engineering constraints into program risk, combined with program management excellence to deliver complex hardware at global scale.

What

You Will Do

You will own programs where the critical path runs through silicon, firmware, and software teams simultaneously. You will translate ambiguity into structure: turning a fleet telemetry signal into a corrective action plan with quantified failure rates, a customer requirement into a new platform milestone with EVT/DVT/PVT gates, or a manufacturing escape into a design change with updated validation criteria. You will drive decisions on program trade-offs — adjusting scope when qualification gates slip, balancing deployment speed against fleet risk, and determining when to accept interim mitigations instead of holding for root-cause fixes.

When a large scale of GPU servers depend on your program landing on time, you are the one ensuring hardware readiness, qualification completeness, and operational handoff happen without gaps.

Why You Will Love It

The world's most advanced frontier models are trained on the platforms you help build. Your programs launch the GPU servers that power the largest AI/ML workloads on the planet. You will see your decisions reflected in fleet reliability metrics within weeks of deployment. The team is small and high-trust — you own programs end to end from concept through production, with direct access to leadership and engineering alike.

The Ideal Candidate

You have deep technical intuition across hardware and software — enough to challenge engineering decisions, not just track them. You thrive in ambiguity, bringing structure to programs where requirements, timelines, and dependencies are still forming. You align priorities across teams in different organizations, and you elevate with data, not noise. You actively mentor and develop others — TPMs and engineers alike.

You contribute to hiring, promotion assessments, and raising the bar for program management practices in your organization.

Key job responsibilities Strategy & Mechanisms
  • Define program strategy, objectives, and success criteria; influence resource allocation and priority decisions across engineering work streams to align with organizational goals
  • Build and own mechanisms for program visibility — defining metrics, dashboards, and review cadences that enable data-driven decisions and early risk detection
  • Streamline delivery processes across teams; identify and eliminate dependencies, redundant gates, or coordination overhead that slow velocity
Requirements & Planning
  • Facilitate requirements gathering with internal customers; develop Technical Requirements Documents (TRDs) covering server specs, rack configurations, PCIe topology, power/cooling topology, and SKU definitions
  • Build program timelines aligned to different phases (Program Initiation, Design, Qualification, Pilot, Post-launch) with critical path analysis, risk identification with new &…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary