Enter a job title or keyword

AIML Sr Manager, Evaluation Data Science & Insights

Apple


Job Location:

Seattle, WA - USA

Monthly Salary: Not provided by the employer
Posted: 22 August 2026 (11 hours ago)
Application Deadline: 19 November 2026
Vacancies: 1 Vacancy

Job Summary

Apple is where individual imaginations gather together committing to the values that lead to great work. Every new product we build service we create or Apple Store experience we deliver is the result of us making each others ideas stronger. That happens because every one of us shares a belief that we can make something wonderful and share it with the world changing lives for the better. Its the diversity of our people and their thinking that inspires the innovation that runs through everything we do. When we bring everybody in we can do the best work of our lives. Here youll do more than join something youll add AIML Evaluation team is looking for a seasoned technical leader to lead a team within our Data Science and Insights organization. The organization leads Evaluation for Apple Intelligence Siri AI and a large portfolio of other billion user facing features in candidates will have strong experience in traditional human evaluation methodology in addition to hands-on experience building and deploying LLM-based autograders and rubrics and using these tools to proactively drive improvements in models and agentic features.

As a Senior Manager on the Data Science and Insights team youll lead a team of data scientists focused on a core pillar of evaluation in close collaboration with teams across the company. Your experience will enable you to thoughtfully balance the various tradeoffs involved in creating successful features that meet Apples high customer expectations for both quality and privacy.

Owning the evaluation roadmap for your area translating the orgs broader evaluation strategy into concrete methods and a team of data scientists recruiting developing and retaining strong technical talent across both hands-on execution across human evaluation and LLM-based autograders and rubrics and ensuring those methods translate into measurable model and agentic feature with Apple Intelligence Siri AI and other SWE product and engineering teams to embed evaluation into product development cycles and turn evaluation results into shipped quality your teams work to senior leadership across AIML and SWE.

6 years of experience in data science and machine learning evaluation including 3 years leading technical teamsnAdvanced degree in a quantitative field such as Statistics Computer Science Machine Learning or similarnDemonstrated track record of running teams of 7 data scientists and/or machine learning engineersnStrong experience in human evaluation methodology for consumer-facing products at scalenHands-on experience building and deploying LLM-based autograders and rubrics and using them to drive proactive improvements in models and agentic featuresnStrong written and verbal communication skills able to communicate effectively with engineers and senior leaders

Experience evaluating large consumer AI products such as conversational assistants search systems or agentic featuresnExperience with logging infrastructure and instrumentation for AI product quality measurementnTrack record of growing senior individual contributors and leads from within your team and where needed recruiting data science and machine learning talent in competitive hiring marketsnFamiliarity with evaluation frameworks for agentic systems and tool-use

Required Experience:

Manager


About Company

Company Logo

Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more

View Profile View Profile