Site Reliability Engineering (SRE) Manager, Apple Maps

Apple


Job Location:

Cupertino, CA - USA

Monthly Salary: Not Disclosed
Posted on: 22 hours ago
Vacancies: 1 Vacancy

Job Summary

Apple Maps and location services are used by hundreds of millions of people every day to navigate the world. Behind every search route and point of interest is a massive distributed serving infrastructure that must be fast reliable and always SRE org is responsible for the availability and automation of some of the most visible and widely used services that power Apple Maps. nnIf you are passionate about reliability at planet scale and are excited to help us build grow manage and deliver infrastructure that scales with Apple Maps this is the opportunity for you!

We are looking for a senior SRE leader to set the strategic direction for our Maps serving infrastructure. This is not just a keep the lights on role this is a leadership position that defines where our infrastructure goes next how our SRE practice evolves and how we build the teams and partnerships to get there. You will work closely with engineering product and operations partners across Apple to shape the roadmap for our serving platform. You will lead an organization of SREs and hold responsibility for the reliability scalability and operational excellence of some of Apples most visible services. We believe AI will fundamentally reshape how SRE is practiced from incident detection and resolution to capacity planning and toil elimination and were looking for a leader who shares that conviction and can drive that transformation across the organization.

Operate with the fundamental principle that reliability is feature number onennDefine and drive the strategic roadmap for Apple Maps serving infrastructure in partnership with engineering and product teamsnnLead and grow an SRE organization setting the bar for technical excellence operational rigor and engineering culturennRepresent the SRE perspective in cross-functional planning translating reliability requirements into architecture decisions and investment prioritiesnnChampion the Engineering in Site Reliability Engineering driving automation platform improvements capacity strategy and systems design beyond reactive incident responsennDefine and execute a clear AI strategy for the SRE organization identifying high-impact opportunities where AI/ML-powered tooling can reduce toil accelerate root cause analysis and improve reliability outcomesnnDrive adoption of AI-assisted tooling (copilots intelligent runbooks LLM-based diagnostics anomaly detection) into day-to-day SRE workflowsnnBuild a culture where engineers actively experiment with AI tools and modern approaches to solve operational problemsnnBuild strong partnerships across Apple negotiating priorities and aligning on shared goals with an Apple-first mindsetnnCommunicate effectively at the executive level presenting strategy trade-offs and progress to senior leadershipnnMentor and develop leaders within your organization creating a culture where people do their best worknnCelebrate wins recognize contributions and make tough calls when needed

10-15 years of experience in SRE or adjacent disciplines (systems engineering infrastructure engineering production engineering) with at least 10 years in senior management rolesnDemonstrated experience leading SRE organizations supporting large-scale user-facing distributed servicesnStrong technical proficiency in Linux fundamentals distributed systems concepts networking and infrastructure at scalenDemonstrated experience applying AI/ML tooling or LLM-based solutions to improve SRE or infrastructure operationsnAbility to read and understand code produced by LLMs and evaluate its suitability for production usenHas defined or is actively executing an AI strategy for a large SRE organizationnProven ability to communicate at the executive level and negotiate across organizational boundaries

Experience with cloud infrastructure (AWS GCP) and Kubernetes at scalenBackground in capacity planning performance engineering or infrastructure architecturenTrack record of driving cultural and process transformation within SRE organizationsnExperience building or deploying AI-powered operational tooling (AIOps intelligent alerting automated diagnostics)nHands-on experience with LLM-based developer/SRE productivity toolsnTrack record of driving AI adoption within engineering teamsnExperience operating services at Apple-scale user volumes

Required Experience:

Manager

Apple Maps and location services are used by hundreds of millions of people every day to navigate the world. Behind every search route and point of interest is a massive distributed serving infrastructure that must be fast reliable and always SRE org is responsible for the availability and automati...

About Company

Company Logo

Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more

View Profile View Profile