Director, Data Center Facilities Operations
xAI
- Location
- Memphis, TN; Southaven, MS
- Work model
- On-Site
- Level
- Staff
- Posted
- 7h ago
About this role
SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.
ABOUT THE ROLE
SpaceX is launching at an incredible tempo out of multiple launch complexes and processing more payloads than ever. The Director of Data Center Facilities Operations is responsible for the strategic leadership, reliability, and operational excellence of all data center facilities infrastructure across an entire SpaceX site. This role owns the full lifecycle of critical infrastructure supporting multiple data centers—including power systems, HVAC/cooling, fire protection, compressed air, plumbing, building structures, and general facilities utilities—ensuring zero unplanned downtime and maximum availability for mission-critical operations.
RESPONSIBILITIES
• Lead and develop a multi-layered organization of managers, supervisors, and technicians across multiple data centers to deliver 24/7/365 reliability with zero unplanned interruptions to employees, customers, and flight-critical systems.
• Set site-wide strategy, multi-year roadmaps, budgeting (CapEx/OpEx), and resource forecasting for all data center facilities infrastructure; drive extreme ownership of total execution, safety culture, and reliability engineering.
• Partner with engineering, IT, operations, and executive leadership to plan, prioritize, resource, and execute large-scale projects that support launch manifests, capacity expansion, and efficiency improvements across the site.
• Establish and continuously improve standardized work instructions, maintenance strategies, and operational processes for all data centers; drive best practices, lessons learned, and continuous improvement across the entire portfolio.
• Conduct ongoing assessment of resource needs, life-cycle costs, and capacity planning for multiple facilities; own long-term capital planning and total cost of ownership for data center infrastructure.
• Oversee all contractors and vendors supporting mechanical, structural, HVAC, electrical, power, cooling, fire/life safety, and water systems across multiple data centers; enforce corporate standards, SLAs, and performance.
• Own site-wide preventative and predictive maintenance programs; develop prioritization frameworks and multi-year strategies for major repairs, upgrades, and replacements to maximize uptime and efficiency (including PUE optimization).
• Define, track, and communicate key performance indicators (safety, uptime/availability, quality, labor productivity, energy efficiency, and operating expenses) to executive leadership and drive accountability across the organization.
• Provide strategic oversight of emergency response and on-call frameworks; ensure robust 24x7 support models and rapid recovery capabilities for all data center facilities.
• Lead talent development, succession planning,