Sr. Linux Engineer DNS and Global Server Load Balancing, Infrastructure Services
Sunnyvale, CA - USA
Job Summary
This is a senior individual-contributor role and a technical anchor for the Edge Services DNS and GSLB systems. The essential functions of the job are to operate and extend distributed Linux infrastructure as code across geographically dispersed sites; to define and maintain the observability alerting and service-level objectives that keep the systems healthy; to build automation that removes operational toil; to plan fleet capacity and hardware lifecycle across globally distributed sites; and to mentor engineers while partnering across software network and product teams. The role includes incident and process management and participation in a shared 24x7 production on-call rotation.
Operates and extends bare-metal and cloud Linux infrastructure across dozens of edge sites and points of presence treating infrastructure as code so that standing up or rebuilding a site is predictable and the reliability correctness and performance of the authoritative DNS and global server load balancing systems including incident response and participation in a 24x7 on-call deep visibility into the systems by defining meaningful service-level objectives alerting and dashboards that span every point of presence and surface problems automation and tooling in Python Go Rust and/or Swift supported by generative AI tooling to remove operational toil and turn manual tasks into reliable self-sustaining fleet capacity growth and hardware lifecycle across geographically distributed sites through smart tooling thoughtful planning and collaboration with leadership and program and grows engineers at every career stage and partners with other senior engineers to raise the bar on architecture tooling and with software network and product engineering teams to deliver large-scale rollouts and keep 24x7 operations running the teams work needs and ideas to partners and leadership through clear collaborative communication.n
Demonstrated experience operating Linux systems in production including software deployment and CI/CD knowledge of networking fundamentals with hands-on experience troubleshooting TCP/UDP and common layer 23 with configuration management or infrastructure-as-code tooling (for example Salt Ansible Puppet or Terraform) and with observability tooling (for example Prometheus and Grafana or equivalents).nProficiency in at least one programming language used for automation and tooling (for example Python Go Rust or Swift).nWillingness to participate in a shared 24x7 on-call degree in Computer Science or a related field or equivalent practical experience.n
Extensive experience operating large-scale infrastructure across multiple data centers or network points of with anycast routing and BGP and a strong understanding of how DNS load balancing and the network layer in Linux internals including kernel networking and in package management and software deployment at fleet defining service-level objectives and indicators (SLOs/SLIs) and driving reliability programs across distributed with secure-by-default operations such as DNSSEC key management change safety and progressive configuration with scale and performance testing disaster recovery and capacity managing hardware lifecycle across edge and regional network record of leading end-to-end projects defining technical roadmaps and driving cross-functional alignment on architecture and best mentoring engineers and leading code reviews.n
Required Experience:
Senior IC
About Company
Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more