Remote job
Technical Program Leader - AI Infrastructure
Job details
About this role
Role overview Lead large-scale AI infrastructure and data center deployment programs from early planning through commissioning, validation, and production handoff. The role combines technical program leadership with hands-on coordination across engineering, facilities, operations, supply chain, contractors, and technology vendors. Work may span multi-megawatt capacity, high-density GPU environments, multiple sites, and complex deployment dependencies.
Responsibilities - Own deployment plans, schedules, milestones, dependencies, critical paths, risks, issues, decisions, and readiness criteria. - Coordinate delivery across site readiness, racks, structured cabling, power, cooling, networking, storage, and accelerator infrastructure. - Lead commissioning, integrated systems testing, burn-in, validation, punch-list closure, and operational readiness activities. - Work with engineering, infrastructure, facilities, operations, OEMs, contractors, and vendors to remove blockers and maintain delivery momentum. - Provide clear updates to technical and executive stakeholders on progress, risks, trade-offs, and required decisions. - Establish repeatable processes that improve the efficiency and reliability of infrastructure deployments at scale.
Requirements - At least 7 years of experience in data center infrastructure, technical deployment, mission-critical facilities, cloud, hyperscale, HPC, AI infrastructure, or a related environment. - Practical understanding of several infrastructure domains, such as rack deployment, cabling, power, cooling, networking, storage, GPU platforms, commissioning, or operational readiness. - Experience leading technical workstreams and coordinating engineering teams, vendors, contractors, and other stakeholders. - Strong knowledge of deployment planning, technical acceptance, defect management, change control, production handoff, and executive reporting. - Ability to engage credibly with engineers, work through complex dependencies, manage vendor accountability, and communicate technical issues as clear actions. - A technical degree is preferred but not required; U.S. work authorization is required and visa sponsorship is unavailable.
Nice to have Experience with accelerator platforms, high-performance networking, AI/HPC storage, liquid cooling, high-density racks, cluster bring-up, performance acceptance, Slurm, Kubernetes, monitoring, or GPU fleet operations. Project, service-management, commissioning, data center, cloud, networking, or platform certifications are also valued.
Benefits and work setup Hybrid work is available in Bellevue, Washington, with three office days per week for candidates within commuting distance. Qualified candidates elsewhere in the U.S. may work fully remotely. Travel of up to 50% may be required.