Performance & Capacity Engineer

Meta

$150K — $180K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, or equivalent experience
  • 8+ years in performance or software engineering and/or relevant optimization data science
  • 8+ years designing and implementing optimization models
  • 4+ years coding experience with languages such as Python, R, Java, etc.
  • Experience with distributed systems at scale
  • Infrastructure operations knowledge
  • Familiarity with cross-functional collaboration

Responsibilities

  • Own capacity planning for infrastructure including Servers, Data Centers, and Network
  • Design and launch software systems to optimize capacity planning
  • Contribute to capacity planning processes and methodologies for executable plans
  • Manage critical escalations in capacity planning
  • Build models for simulation and optimization of demand and supply
  • Collaborate with cross-functional teams to drive infrastructure decisions
  • Balance operational demands with long-term project goals

Benefits

  • Opportunities to work with innovative Meta products
  • Collaboration with diverse engineering and business teams
  • Fostering of AI and emerging technology skills
  • Impacting global infrastructure strategies
  • Involvement in scaling massive infrastructure operations
Full Job Description
Performance & Capacity Engineer to join the Capacity Planning and Optimization Engineering (CPOE) team to focus on site-wide performance and capacity optimization at the intersection of all Meta products and services, and all physical infrastructure (Servers, Data Centers, Network). This role will focus on creating and optimizing capacity plans at various altitudes to balance capacity, power, and cost, in an environment of continual product introduction, directly managing billions of investment annually. We do this primarily by building software and mathematical optimization models, not manual planning. This role will be uniquely positioned to optimize these capacity plans and scalably manage exceptions to drive company-level infrastructure decisions. This role requires collaboration across many engineering and business teams, with the opportunity to work with a variety of engineering and business teams to power the rapid growth of Meta's products. This role will help to ensure optimal operation of our infrastructure from both a cost and technology perspective, with millions of servers, Gigawatts of data center capacity, and current and emerging technology including AI, Metaverse, etc. Help build and scale Meta's global infrastructure fleet!

Responsibilities

Own infrastructure capacity planning for Meta: including Servers, Data Centers, Network
• Design, implement and launch software systems to improve capacity planning efficiency and quality, partnering with software engineers
• Contribute to end to end capacity planning processes, methodologies, and data to deliver executable and optimized plans
• Manage and resolve critical escalations and exceptions in all areas of the capacity planning
• Build mathematical models to perform simulation and optimization studies of demand and supply projections, scenario planning, and feasibility analysis while balancing various constraints
• Work cross-functionally to define problem statements, collect data, build analytical models and make recommendations to drive change and optimization to inform company-level infrastructure decisions
• Partner across Infra: such as platform teams, operations, networking planning, data center planning as well as Product and Finance teams to find the most optimal ways to scale our Infrastructure
• Effectively navigate complex tradeoffs and relationships to balance solving for team, cross-functional partner / stakeholders, and Meta company priorities. Balance the need to "keep things running" with longer-term, high-impact projects

Minimum Qualifications
• Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
• 8+ years of experience in performance or software engineering and/or optimization pertinent data science or equivalent practical experience
• 8+ years of experience in designing and implementing models and optimization algorithms
• 4+ years of experience in coding/scripting languages such as Python, R, Java, C, C++, PHP
• Experience working with distributed systems at scale
• Experience in infrastructure operations and technical infrastructure knowledge
• Experience working with cross-functional teams
• Experience optimizing complex systems, working with large datasets, and driving business impact

Preferred Qualifications
• Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
• Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
• Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
• Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
• Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
• Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
• Experience in multi-phase/multi-year system/software roadmap development
• Experience in public or private cloud capacity planning and optimization

Similar Jobs

More Jobs at Meta

More Information Technology Jobs

Find similar Performance & Capacity Engineer jobs: