Meta is seeking a Technical Program Manager (TPM) with capacity planning and delivery experience at scale. This position will work with cross-functional teams to drive the planning, allocation, and delivery of compute, storage, and network capacity across Meta's global infrastructure. This role requires engagement with internal engineering teams, data scientists, and stakeholders to ensure capacity meets the demands of Meta's products and services. Meta's Infrastructure Engineering organization is responsible for the growth, management, and 24x7 upkeep of all Meta's products and services.
Responsibilities
- Lead technical program management of capacity planning and delivery for Meta Infrastructure in a matrix organization covering compute, storage, GPU, and network capacity
- Own overall program success spanning end-to-end capacity planning cycles, from demand forecasting through allocation and delivery to production workloads at scale
- Develop and manage capacity programs including defining scope, requirements, delivery milestones, schedules, and deliverables with engineering teams, partners, and stakeholders
- Drive capacity forecasting and modeling efforts, collaborating with data science and engineering teams to anticipate infrastructure needs
- Provide hands-on program management during capacity planning, allocation, deployment, and optimization phases
- Perform risk assessment, risk mitigation, and change management on capacity programs
- Drive process improvements across capacity planning, supply chain, and delivery workflows
- Coordinate cross-functional capacity reviews and communicate status, risks, and tradeoffs to leadership
Minimum Qualifications
- 12+ years of experience in technical program management, infrastructure engineering, or a related technical discipline
- 8+ years of experience in Technical Program Management
- Experience leading large-scale, cross-functional programs with significant organizational impact
- Experience driving strategic technical initiatives and influencing roadmaps across multiple teams
- Experience with executive-level communication, including presenting to senior leadership
- Experience mentoring and developing other technical program managers
- Experience navigating ambiguity and defining structure for complex, multi-year programs
- Experience with capacity planning, infrastructure systems, or related domains Experience with data center architecture and deployment
- Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
- Experience supporting AI/ML infrastructure capacity needs
- Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
- Experience with capacity modeling, demand forecasting, or resource optimization
- Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
- Experience with supply chain planning and vendor coordination