Skip to main contentA logo with &quat;the muse&quat; in dark blue text.

AI Inference Platform Engineer

Yesterday Seattle, WA

We are looking for an engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform. This role will focus on developing sophisticated performance benchmarking systems, capacity projection models, and data analysis pipelines that directly inform our AI infrastructure teams and capacity planners. You'll work at the intersection of AI systems performance, distributed infrastructure, and software engineering to help the team make data-driven decisions about scaling and optimizing our inference platform.

Description

At Apple, we believe the future of AI is defined not just by models, but by the infrastructure that powers them. Our AI inference platform sits at the heart of products and experiences used by hundreds of millions of people worldwide, and we are building the systems that ensure it scales reliably, efficiently, and intelligently.

As part of our next-generation datacenter engineering team, you will play a critical role in shaping how we understand, measure, and grow our AI infrastructure. You will design and build the tooling and analysis systems that give our engineers and capacity planners a clear, real-time picture of performance across our fleet. Your work will directly influence how we invest in hardware, how we detect regressions before they reach production, and how we forecast capacity needs months in advance.

Want more jobs like this?

Get Data and Analytics jobs in Seattle, WA delivered to your inbox every week.

Job alert subscription


This is a high-impact, cross-functional role for an engineer who is energized by complexity, thrives on turning raw data into actionable insight, and wants to work on problems that matter at massive scale.

Responsibilities:

Design and build automations to evaluate AI inference performance across hardware generations and configurations.

Develop tooling to surface performance trends, regressions, and insights to infrastructure and planning teams.

Build projection and forecasting models to support long-term capacity planning decisions.

Analyze performance and utilization data to identify bottlenecks, trends, and optimization opportunities.

Partner with AI infrastructure engineers, hardware teams, and capacity planners to deliver critical data and tooling.

Create and enhance performance analysis workflows to increase team velocity and data reliability.

Continuously improve the accuracy, coverage, and usability of performance measurement and analysis systems.

Preferred Qualifications

Experience with performance benchmarking and methodologies for AI/ML inference systems.

Familiarity with capacity planning and forecasting/projection models for large-scale infrastructure.

Experience with GPU profiling and observability tools (e.g., Nsight, other vendor-specific profilers).

Experience with data visualization and reporting tools/frameworks for surfacing performance trends to stakeholders.

Familiarity with ML serving frameworks and runtimes (e.g., Triton, TensorRT-LLM, vLLM, or similar).

Experience with CI/CD and workflow orchestration tools for building automated performance analysis pipelines.

Knowledge of cluster schedulers and orchestration platforms (e.g., Kubernetes).

Experience with metrics and logging tools (e.g., Prometheus, Grafana, Splunk).

Minimum Qualifications

BS or MS in Computer Science or related technical field.

Solid understanding of AI/ML inference architecture and the performance characteristics of serving systems.

Experience with performance and infrastructure engineering in distributed systems.

Proficiency in Python, Go, C++, or other programming languages.

Experience with automation engineering, tooling, and data pipelines to support engineering workflows.

Strong knowledge of GPU/accelerator architecture as it relates to AI workloads.

Practical statistical knowledge applicable to performance analysis and forecasting.

Excellent communication skills and ability to turn data into clear guidance for infrastructure teams and capacity planners.

Pay & Benefits

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $175,000 and $308,500, and your base pay will depend on your skills, qualifications, experience, and location.

Apple employees also have the opportunity to become an Apple shareholder through participation in Apple's discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple's Employee Stock Purchase Plan. You'll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses - including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits

Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program.

Client-provided location(s): Seattle, WA
Job ID: apple-200674824-3337
Employment Type: OTHER
Posted: 2026-08-14T21:14:40

Perks and Benefits

  • Health and Wellness

    • Parental Benefits

      • Work Flexibility

        • Office Life and Perks

          • Vacation and Time Off

            • Financial and Retirement

              • Professional Development

                • Diversity and Inclusion

                  Company Videos

                  Hear directly from employees about what it is like to work at Apple.