at Apple
Location
San Francisco Bay Area, United States of America
Type
full time
Posted
1 weeks ago
Market range · company + function + seniority
p25 · target · p75 · n=53
Tailor your résumé to this role in 30 seconds.
Free account · ATS keyword check · per-job bullet rewrite by Claude.
This role focuses on leading a team working on developing, carrying-out, interpreting, and communicating pre- and post-ship evaluations of the safety of Apple Intelligence features. This team is also responsible for producing safety evaluations that uphold Apple’s Responsible AI values requires thoughtful data sampling, creation, and curation for evaluation datasets; high quality, detailed annotations and careful auto-grading to assess feature performance; and mindful analysis to understand what the evaluation means for the user experience.
Proven technical leadership with 5+ years of team management or leadership experience
Demonstrated expertise in ML production deployment lifecycle, datasets, identifying data needs, and working on creative solutions, scaling and expanding data coverage through human and synthetic generation methods.
Provide technical direction and expertise to team-wide initiatives in safety auto-grading.
Use and implement data pipelines, and collaborate cross-functionally to execute end-to-end safety evaluations.
Work with highly-sensitive content with exposure to offensive and controversial content.
MS, or PhD in Computer Science, Machine Learning, Statistics, or related fields; or an equivalent qualification acquired through other avenues.
Experience working with generative models for evaluation and/or product development, and up-to-date knowledge of common challenges and failures.
Strong engineering skills and experience in writing production-quality code in Python.
Deep experience in foundation model-based AI programming (i.e.: using DSPy for optimizing foundation model prompts, for example) and a drive to innovate in this space.
Experience working with noisy, crowd-based data labels and human evaluations.
Publication record in relevant conferences (e.g., NeurIPS, ICML, ICLR, EMNLP, etc.)
Experience working on Responsible AI and AI Safety
Strong organizational and operational skills working with large, multi-functional, and diverse teams.
Curiosity about fairness and bias in generative AI systems, and a strong desire to help make the technology more equitable.
Apple's Responsible AI and Safety team focuses on innovative technologies, methodologies, and research to enable fantastic user experiences and to push the frontier of machine learning. Our team is looking to hire a leader with a strong track record in Applied Research, who is passionate about ML and foundation models with a focus on responsibility, fairness, and safety. In this role, you will lead the research and application of ML methods for technologies that power breakthrough user experiences while upholding Apple's values, privacy, and quality standards.
This posting is not for a specific job opening and by submitting your resume you are expressing interest in being contacted about this type of role at Apple in the future.Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant
At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.
Learn about accessibility in Apple’s workplace
Learn about reasonable accommodations for job applicants
Apple accepts applications to this posting on an ongoing basis.
More open roles at Apple
Hiring velocity, headcount trend, and every open posting on one page.
Open postings ranked by description similarity — useful if this role isn't quite right.