SRE, London
London, KY - USA
Job Summary
FoundationDB infrastructure is BIG. Operating at our scale across multiple geographically dispersed data centers and servicing hundreds of millions of users presents unique challenges. As an SRE at Apple youll need to solve these problems using data teamwork and your own expertise. SREs at Apple own the full infrastructure stack; from device driver performance debugging to content delivery network traffic management our responsibilities are both broad and runs its systems on Linux. We run a mix of open source vendor licensed and internally developed tools to perform functions such as system configuration management provisioning software deployment logging and monitoring. Youll learn these tools and have opportunities to improve them. Our team is collaborative; we work closely with the development teams we support to deliver the best results for Apple. We think critically and strive to balance the best solution with the need to get things done for each engineering challenge we face. Good ideas are heard and results are SRE is a small team with huge scale. We serve as the database for much of CloudKits use cases including Mail Contacts and Keychain. We serve hundreds of millions of customers every day and are a fundamental piece of the Apple device
Our team is responsible for the provisioning managing and monitoring of FoundationDB in production across multiple regions and control planes (bare-metal AWS and Kubernetes). We develop much of our own automation in Java and Go including our open source Kubernetes Operator ( We work closely with our dev partners to develop a robust and scalable database often engaging in projects as a single team.
Strong sense of ownership and integrity demonstrated through clear communication and collaborationnExperience in managing and scaling distributed systems in a public private or hybrid cloud environmentnThe ability to design author and release code in languages like (but not limited to) Go Java or PythonnAcute drive to automate manual operations and to improve them through repeated iterationnUnderstanding of the Linux Operating System standard networking protocols and componentsnHands-on experience managing large numbers of diverse systems with configuration management or software delivery platforms (such as Puppet Chef Ansible and Spinnaker)nExperience with deploying supporting and monitoring new and existing services platforms and application stacksnExcellent troubleshooting and problem solving skillsnExperience with scale testing disaster recovery and capacity planningnFamiliarity with microservices architecture and container orchestration with Kubernetes
Hands-on experience managing large numbers of diverse systems with configuration management or software delivery platforms (such as Puppet Chef Ansible and Spinnaker)nExperience with deploying supporting and monitoring new and existing services platforms and application stacksnExcellent troubleshooting and problem solving skillsnExperience with scale testing disaster recovery and capacity planningnFamiliarity with microservices architecture and container orchestration with Kubernetes
About Company
Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more