Description
Summary
Apple's Platform Acceleration & Compute Efficiency (PACE) is a high-leverage team operating at the intersection of our ML organizations, underlying compute infrastructure, and core platform tooling. Our mission is to empower Apple's software engineering teams with efficient, scalable compute. By driving out operational friction and optimizing the broader machine learning ecosystem, we directly accelerate the pace of development for our Software organization.
Foundation models are central to Apple's user experiences and maximizing the efficiency of our ML compute is paramount. Efficiency sits at the center of this role, ensuring that Apple's models run as fast, reliably, and cost-effectively as possible. In this role you will tackle optimization challenges, from maximizing hardware utilization across GPUs, TPUs, and custom Apple Silicon, to shaping workload scheduling and capacity allocation for large model serving.
In addition to overseeing accelerated compute, this role will focus on policy setting and enforcement of GenAI access and the related token consumption. Finally, this role will be responsible for overseeing traditional 1P vs 3P managed cloud services (e.g., non-accelerated compute, storage, etc.) as well as any process-related initiatives to improve the efficiency and effectiveness of the overall program.
We are seeking a seasoned Engineering Program Manager who can apply structure, discipline, and accountability to this growing area of complexity for the Company. This role is a true partnership with not only the PACE team, but also the broader Infrastructure and Planning organization at large. You will support the execution of the Governance Platform's feature roadmap and its day-to-day operating model which includes monthly / quarterly executive roll-ups, partnering with the team leads to keep their respective roadmaps current and presentation-ready for leadership.
Description
* Support the operational rollout of the hub-and-spoke support model: Governance DRI onboarding, and the transition of manual processes into scalable systems.
* Program manage GenAI policies at Apple that can be established and enforced using sophisticated web-tooling.
* Program manage 1P vs 3P monthly forecast submissions and partner with technical leads on any related process efficiencies or optimization initiatives.
* Partner with the Compute Efficiency lead and Intelligent Automation lead to keep their roadmaps current and accurately reflected in leadership reviews.
Minimum Qualifications
Bachelor's Degree
8+ years of relevant work experience managing SW, ML, or compute capacity programs
Demonstrated experience scaling operational structure (support tiers, escalation processes, onboarding) for a fast-growing team or program
Experience co-driving feature delivery with an engineering or platform lead, translating architecture/systems design into a shippable roadmap
Proven track record owning recurring executive or leadership review cadences (QBRs, MBRs, or equivalent) end to end - content, prep, and delivery
Ability to operate with a high level of ambiguity and complexity in a founding, ground-up program.
Strong interpersonal skills with a proven track record collaborating across diverse, cross-functional teams
Preferred Qualifications
Outstanding verbal and written communication skills for clearly presenting findings to key partners and executives
Familiarity with cloud/ML compute infrastructure concepts (GPUs, TPUs, capacity, utilization)
Experience partnering with Governance, FinOps, or platform teams in a large-scale infrastructure environment
Background in change management or process design during periods of rapid team/program growth
Outstanding attention to detail with a focus on accuracy and precision in analysis and reporting
MBA preferred for important connection of engineering complexity with business and financial impact