Principal Technical Infrastructure Program Manager, AWS DC Acqn&Construction
Amazon
- Location
- US, WA, Seattle
- Employment
- Full Time
- Work model
- On-Site
- Level
- Principal
- Posted
- Aug 27, 2026
Skills
About this role
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help. You’ll join a diverse team of technical program managers, design engineers, construction managers, network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion. The ML Capacity Delivery Team (MLZ) is seeking a Principal Technical Infrastructure Program Manager to define and drive the technical vision, strategy, and roadmap for the systems and tools that enable ML capacity delivery at global scale. In this role, you will serve as the single-threaded technical leader responsible for building, evolving, and scaling the platforms and automation that underpin how we plan, track, and deliver ML infrastructure; from demand signal through rack-level installation. You will work at the intersection of infrastructure delivery, data engineering, and tooling while partnering with systems engineers, software development teams, Business Intelligence Engineers (BIE), operations leaders, and cross-functional stakeholders to translate complex operational requirements into scalable, reliable systems. The ideal candidate is a seasoned technical leader who thrives in ambiguity, can articulate a compelling 1–3 year technical roadmap, and has a track record of driving cross-organizational alignment to deliver platforms that fundamentally improve how teams operate at scale. Key job responsibilities • Own the technical vision and multi-year roadmap (1–3 years) for ML Capacity Delivery systems and tools, aligning investments to business priorities and capacity delivery goals across core and Gen AI platforms. • Drive cross-functional alignment across systems engineering, software development, BIE, operations, and planning teams to define requirements, prioritize capabilities, and deliver integrated tooling solutions. • Architect and deliver scalable platforms that automate and optimize capacity delivery workflows; including demand planning, build tracking, supply chain visibility, and operational reporting. • Serve as the technical authority for systems and tools within MLZ, influencing Director/VP-level leadership on investment decisions, build-vs-buy trade-offs, and platform strategy through data-driven business cases. • Identify and eliminate operational friction by mapping end-to-end workflows, quantifying inefficiencies, and delivering automation that reduces manual effort and accelerates delivery velocity. • Establish engineering excellence; define standards for system reliability, data quality, observability, and documentation that enable teams to build and operate with confidence at scale. • Force-multiply across the organization by creating reusable frameworks, mechanisms, and best practices that elevate the effectiveness of multiple teams and programs simultaneously. • Drive clarity in highly ambiguous environments where the business strategy, architectural approach, and problem definition may not yet exist. While decomposing complex challenges into actionable execution plans. • Build and maintain strategic partnerships with internal platform teams, data engineering organizations, and tooling providers to ensure MLZ systems integrate seamlessly with the