GPU Deployment Execution Program Manager
Microsoft
- Location
- United States, Washington, Redmond; United States, Texas, San Antonio; United States, Georgia, Atlanta; United States, Arizona, Phoenix; United States, District of Columbia, Washington D.C.
- Work model
- On-Site
- Level
- Staff
- H-1B history
- 2,066 approvals (FY2023)
- Posted
- 2h ago
Skills
About this role
Overview
Microsoft’s Cloud Operations & Innovation (CO+I) powers our cloud services and supports over 1 billion customers and 20 million businesses in 90+ countries. With over 100 data centers and 1 million servers worldwide, CO+I is focused on delivering secure, reliable, and sustainable infrastructure. We are committed to fostering an inclusive work environment where all employees thrive—and we need you as a GPU Deployment Execution Program Manager to help us lead the next wave of infrastructure evolution. As a key member of the team, you’ll help drive Microsoft’s GPU deployments across key metros, supporting services like Azure, Xbox, Office 365, and more. You’ll work in a fast-paced, cross-functional environment to deliver high-impact results and advance Microsoft’s mission to empower every person and organization on the planet to achieve more.
Responsibilities
Participate in GPU-related planning forums including capacity expansion, retrofit initiatives, and deployment strategy to align execution with business and technical requirements. Lead the development, execution, and communication of GPU deployment project plans across metro data centers, ensuring alignment with operational timelines and infrastructure readiness. Drive execution of GPU deployment activities the Americas region, streamlining processes to ensure timely delivery of capacity and services, ensuring all roles are prepared to support GPU growth. Serve as the primary liaison between metro leadership and operations, escalating risks and representing deployment interests within the regional program framework. Partner closely with Supply Chain, Engineering Groups, and Networking teams to generate execution schedule and troubleshoot and resolve technical issues throughout the deployment lifecycle. Develop and maintain risk management plans to ensure continuity and resilience during deployment and sustainment operations. Document lessons learned and implement best practices in GPU capacity planning to improve accuracy, efficiency, and long-term scalability. Embody our culture and values .
Qualifications
Required Qualifications: High School Qualification or equivalent AND 10+ years experience supporting IT equipment or related technology or delivering server and network deployment projects in large-scale environments OR equivalent experience Other Requirements: Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings: Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.
Preferred Qualifications
Bachelor's or Technical College Degree in Computer Science, Math, Telecommunications, Electrical/Mechanical Engineering, Supply Chain Management or related field AND 15+ years experience in critical environment infrastructures (e.g., UPS, Generator, AHU), or working in physical IT infrastructures (e.g., Servers, SANs, Networking, Capacity, DC Rack/Enclosures, structured cabling) OR High School Qualification or equivalent AND 17+ years experience in critical environment infrastructures (e.g., UPS, Generator, AHU), or working in physical IT infrastructures (e.g., Servers, SANs, Networking, Capacity, DC Rack/Enclosures, structured cabling) OR equivalent experience Applicable certifications: Microsoft, Network Certifications, CCNA Certifications, ITIL v3 Foundation, Microsoft Operations Framework (MOF) Certifications 5+ years in IT Operations leadership (preferably in data centers) 10+ years working with physical IT infrastructure (servers, networking, cabling, etc.) Demonstrated knowledge of physical IT infrastructures (e.g. Servers, SANs, Networking, Capacity, DC Rack/Enclosures, structured cabling, etc.) Data Center Operations Management IC6 - The typical base