at Apple
Location
Seattle, United States of America
Compensation
$175k–$309k USD
Type
full time
Posted
1 months ago
Market range · company + function + seniority
p25 · target · p75 · n=800
Posted $309k · above the band
Posting health
Aging · 60Tailor your résumé to this role in 30 seconds.
Free account · ATS keyword check · per-job bullet rewrite by Claude.
At Apple, we believe the future of AI is defined not just by models, but by the infrastructure that powers them. Our AI inference platform sits at the heart of products and experiences used by hundreds of millions of people worldwide, and we are building the systems that ensure it scales reliably, efficiently, and intelligently.
As part of our next-generation datacenter engineering team, you will play a critical role in shaping how we understand, measure, and grow our AI infrastructure. You will design and build the tooling and analysis systems that give our engineers and capacity planners a clear, real-time picture of performance across our fleet. Your work will directly influence how we invest in hardware, how we detect regressions before they reach production, and how we forecast capacity needs months in advance.
This is a high-impact, cross-functional role for an engineer who is energized by complexity, thrives on turning raw data into actionable insight, and wants to work on problems that matter at massive scale.
Design and build automations to evaluate AI inference performance across hardware generations and configurations.
Develop tooling to surface performance trends, regressions, and insights to infrastructure and planning teams.
Build projection and forecasting models to support long-term capacity planning decisions.
Analyze performance and utilization data to identify bottlenecks, trends, and optimization opportunities.
Partner with AI infrastructure engineers, hardware teams, and capacity planners to deliver critical data and tooling.
Create and enhance performance analysis workflows to increase team velocity and data reliability.
Continuously improve the accuracy, coverage, and usability of performance measurement and analysis systems.
BS or MS in Computer Science or related technical field.
Solid understanding of AI/ML inference architecture and the performance characteristics of serving systems.
7 or more years of experience with performance and infrastructure engineering in distributed systems.
7 years of experience coding in Python, Go, C++, or other programming languages.
Experience with automation engineering, tooling, and data pipelines to support engineering workflows.
Strong knowledge of GPU/accelerator architecture as it relates to AI workloads.
Practical statistical knowledge applicable to performance analysis and forecasting.
Excellent communication skills and ability to turn data into clear guidance for infrastructure teams and capacity planners.
Experience with performance benchmarking and methodologies for AI/ML inference systems.
Familiarity with capacity planning and forecasting/projection models for large-scale infrastructure.
Experience with GPU profiling and observability tools (e.g., Nsight, other vendor-specific profilers).
Experience with data visualization and reporting tools/frameworks for surfacing performance trends to stakeholders.
Familiarity with ML serving frameworks and runtimes (e.g., Triton, TensorRT-LLM, vLLM, or similar).
Experience with CI/CD and workflow orchestration tools for building automated performance analysis pipelines.
Knowledge of cluster schedulers and orchestration platforms (e.g., Kubernetes).
Experience with metrics and logging tools (e.g., Prometheus, Grafana, Splunk).
We are looking for senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform. This role will focus on developing sophisticated performance benchmarking systems, capacity projection models, and data analysis pipelines that directly inform our AI infrastructure teams and capacity planners. You'll work at the intersection of AI systems performance, distributed infrastructure, and software engineering to help the team make data-driven decisions about scaling and optimizing our inference platform.
At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $175,000 and $308,500, and your base pay will depend on your skills, qualifications, experience, and location.Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant
At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.
Learn about accessibility in Apple’s workplace
Learn about reasonable accommodations for job applicants
Apple accepts applications to this posting on an ongoing basis.
More open roles at Apple
Hiring velocity, headcount trend, and every open posting on one page.
Open postings ranked by description similarity — useful if this role isn't quite right.