At Apple we believe that innovation flourishes in an environment where ideas are challenged collaboration is encouraged and technology is pushed to its limits. This environment is only possible when diverse minds come together bringing unique perspectives and experiences. Our people and their ideas inspire innovation in everything we do. Imagine what you could accomplish here! Join Apple and help us make the world a better a principal contributor in our Apple Data Platform SRE organization you will apply SRE principles as you mentor and partner with our engineers and partner teams ensuring petabyte-scale analytics infrastructure runs reliably and efficiently. This role focuses on managing bare-metal and cloud based infrastructure levering and extending our infrastructure-as-code based tooling analyzing and optimizing performance helping to plan and execute long term fleet management logistics capacity planning and ultimately maintaining operational excellence across distributed data platforms that power analytics across Apple. This role includes production on-call responsibilities.
Apple Service Engineering (ASE) teams build and scale the platforms and infrastructure behind many of Apples services (such as iCloud iTunes Siri and Maps). We are the foundation on which Apples software developers build the products that our customers love. We are looking for a passionate and dedicated Senior Site Reliability Engineer to provide technical leadership on our team to help ensure our customers have the highest quality Apple Services experience. The Apple Data Platform (ADP) Compute SRE team is responsible for the core infrastructure including our legacy bare-metal platforms and modern cloud based infrastructure stack. We partner with both peer SRE teams and several of our world-class software and product engineering teams to support infrastructure reliability multi-year parallel migrations for Apple properties as well as the automation tooling incident and process management necessary to ensure smooth 24x7 operations for ADP customers.
Principal/Lead SRE for ADP Compute SRE team mentoring other ICs and partnering with other senior engineers to ensure service architecture tooling design and implementations are of the highest efficiency improvements in technical operations and training partner teams leading by infrastructure fleet capacity growth and hardware lifecycle with innovative tooling effective planning and proactive check-ins with our leadership and program management technical leadership for Hadoop and Kubernetes infrastructure tooling (including infrastructure-as-code) and in Python and Golang supported by Generative AI tooling to accelerate development of mission critical automation and collaboration and presentation skills to effectively communicate ideas and represent the deliverables and needs of the SRE team with ASE on-call and incident management responsibilities.
12 years of experience in Site Reliability Engineering managing infrastructure and services at scalen5 years of experience in management or technical leadership rolesnHistory of end-to-end project management and deliverynDemonstrable programming skills to both develop software/tools and lead code reviewsnExperience managing Hadoop and Kubernetes infrastructure and related services or equivalent experiencenAdvanced knowledge of Linux Networking and Containers
15 YoE in SRE or related work managing infrastructure at scalenExperience with scale testing disaster recovery and capacity planningnAbility to define the technical roadmap for infrastructure and drive cross-functional alignment on architectural standards and best practices
Required Experience:
Senior IC
At Apple we believe that innovation flourishes in an environment where ideas are challenged collaboration is encouraged and technology is pushed to its limits. This environment is only possible when diverse minds come together bringing unique perspectives and experiences. Our people and their ideas ...
At Apple we believe that innovation flourishes in an environment where ideas are challenged collaboration is encouraged and technology is pushed to its limits. This environment is only possible when diverse minds come together bringing unique perspectives and experiences. Our people and their ideas inspire innovation in everything we do. Imagine what you could accomplish here! Join Apple and help us make the world a better a principal contributor in our Apple Data Platform SRE organization you will apply SRE principles as you mentor and partner with our engineers and partner teams ensuring petabyte-scale analytics infrastructure runs reliably and efficiently. This role focuses on managing bare-metal and cloud based infrastructure levering and extending our infrastructure-as-code based tooling analyzing and optimizing performance helping to plan and execute long term fleet management logistics capacity planning and ultimately maintaining operational excellence across distributed data platforms that power analytics across Apple. This role includes production on-call responsibilities.
Apple Service Engineering (ASE) teams build and scale the platforms and infrastructure behind many of Apples services (such as iCloud iTunes Siri and Maps). We are the foundation on which Apples software developers build the products that our customers love. We are looking for a passionate and dedicated Senior Site Reliability Engineer to provide technical leadership on our team to help ensure our customers have the highest quality Apple Services experience. The Apple Data Platform (ADP) Compute SRE team is responsible for the core infrastructure including our legacy bare-metal platforms and modern cloud based infrastructure stack. We partner with both peer SRE teams and several of our world-class software and product engineering teams to support infrastructure reliability multi-year parallel migrations for Apple properties as well as the automation tooling incident and process management necessary to ensure smooth 24x7 operations for ADP customers.
Principal/Lead SRE for ADP Compute SRE team mentoring other ICs and partnering with other senior engineers to ensure service architecture tooling design and implementations are of the highest efficiency improvements in technical operations and training partner teams leading by infrastructure fleet capacity growth and hardware lifecycle with innovative tooling effective planning and proactive check-ins with our leadership and program management technical leadership for Hadoop and Kubernetes infrastructure tooling (including infrastructure-as-code) and in Python and Golang supported by Generative AI tooling to accelerate development of mission critical automation and collaboration and presentation skills to effectively communicate ideas and represent the deliverables and needs of the SRE team with ASE on-call and incident management responsibilities.
12 years of experience in Site Reliability Engineering managing infrastructure and services at scalen5 years of experience in management or technical leadership rolesnHistory of end-to-end project management and deliverynDemonstrable programming skills to both develop software/tools and lead code reviewsnExperience managing Hadoop and Kubernetes infrastructure and related services or equivalent experiencenAdvanced knowledge of Linux Networking and Containers
15 YoE in SRE or related work managing infrastructure at scalenExperience with scale testing disaster recovery and capacity planningnAbility to define the technical roadmap for infrastructure and drive cross-functional alignment on architectural standards and best practices
Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar
... View more