Site Reliability Engineer, London
London, KY - USA
Job Summary
Apple Services Engineering infrastructure is BIG. Operating at our scale across multiple geographically dispersed data centers and servicing hundreds of millions of users presents unique challenges. As an SRE at Apple youll need to solve these problems using data teamwork and your own expertise. SREs at Apple own the full infrastructure stack; from device driver performance debugging to content delivery network traffic management our responsibilities are both broad and runs the majority of its systems on Linux. We run a mix of open source vendor licensed and internally developed tools to perform functions such as system configuration management provisioning software deployment logging and monitoring. Youll learn these tools and have opportunities to improve them. Our team is collaborative; we work closely with the development teams we support to deliver the best results for Apple. We think critically and strive to balance the best solution with the need to get things done for each engineering challenge we face. Good ideas are heard and results are rewarded. Culturally we believe in a close partnership with our development teams and aim to design u0026 build new services together. Were passionate about software and automation in SRE and develop a variety of tooling and infrastructure. Our services run on mixed u0026 hybrid platforms.n
Create outstanding customer experience and help developers write better code fasternTroubleshoot complex distributed systems running on both bare metal and hypervisorsnParticipate in code reviews and provide helpful and precise feedbacknConstantly evaluate and improve our own delivery processes within the teamnEvolve critical foundational systems to provide next generation features at scale
Advanced experience with programming languages (Go Python Ruby Bash) and a passion for designing and building reliable systemsnStrong sense of ownership and integrity demonstrated through clear communication and collaboration with a deep systems and infrastructure knowledgenAdvanced knowledge and hands-on experience with source code and artifact management systems CI/CD infrastructure (GitHub Artifactory Jenkins)nAutomation advocate - you truly believe in removing operation load with softwarenUnderstanding of the Linux operating system standard networking protocols and components
Experience in managing and scaling distributed systems in a public private or hybrid cloud environmentnHands-on experience managing large numbers of diverse systems with configuration management or software delivery platforms (Puppet Ansible)nExperience with deploying supporting and monitoring new and existing services platforms and application stacks (Grafana Splunk)nExcellent troubleshooting and problem solving skills both with and without AI assistance
Required Experience:
IC
About Company
Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more