Manager, Cloud Engineering AA, Remote
Los Angeles, CA - USA
Job Summary
The Manager Cloud Engineering will continuously collaborate with key stakeholders across the business to solve critical technical problems for Allergan Aesthetics. You will be a hands-on leader of the foundational engineering team driving our system/service reliability scale and developer this dual-focus leadership role you will bridge the gap between keeping our production environment highly resilient and ensuring our internal software developers have friction-free high-speed tools and pipelines. You will lead a talented team of engineers who build scalable infrastructure standardize developer tooling reduce toil and foster a culture of operational excellence across the entire engineering organization.
People & Team Leadership:
- Manage lead hire develop and mentor a team of Platform Engineers to deliver high-quality infrastructure and developer solutions that are secure resilient and scalable.
- Foster a collaborative and inclusive team culture encouraging knowledge sharing and continuous learning.
- Create and recommend professional learning paths for direct reports
- Conceptualize and execute technical approaches for emerging systems and architecture.
- Implement through the team the high bar of quality and standards for production reliability and the developer experience
- Act as the key point of contact between your team and other stakeholders and partner teams
Site Reliability & Platform Engineering:
- Lead the adoption and maturation of SRE practices across the organization partnering with service owners to improve production reliability.
- Define and champion SLO/SLI frameworks and error budget-based decision making to drive prioritization between reliability and feature work.
- Lead the strategy for internal developer platform capabilities influence self-service infrastructure developer tooling and deployment patterns that improve velocity and reduce toil for product teams.
- Responsible for team capacity planning and staffing allocation activities to ensure sufficient staffing for current and future needs.
Incident Management & Response:
- Plan staff and support on-call rotation.
- Manage robust incident response and resolution capture corrective and preventative actions and continuous improvements
Qualifications :
- Bachelors degree in Computer Science Information Technology or a related field.
- Experience (1-3 years) in a leadership role managing cloud/platform engineering teams with 6 years of hands-on experience in SRE DevOps Infrastructure and Platform Engineering.
- Experience building and operating internal developer platforms that improve engineering productivity standardize delivery patterns and reduce operational toil.
- Deep background in modern cloud environments (AWS GCP or Azure) container orchestration (Kubernetes) Infrastructure as Code (Terraform Pulumi) and CI/CD tools (GitHub Actions ArgoCD GitLab CI).
- System Design: Proven track record in designing scaling and maintaining distributed systems with high availability requirements.
- Observability: Experience implementing comprehensive monitoring logging tracing and alerting frameworks (Datadog Prometheus Grafana OpenTelemetry)
- Experience surveying developer teams to identify friction points prioritize platform improvements and translate feedback into self-service capabilities that improve developer productivity.
- Deep experience with platform engineering and SRE practices and frameworks
- Programming proficiency of various languages (e.g. Java C# Python Go Rust)
- Knowledge of secure SDLC practices and methodologies.
- Excellent communication and interpersonal skills with the ability to effectively collaborate with cross-functional teams and senior stakeholders.
- Relevant cloud certifications (e.g. AWS Google Cloud Azure) are a plus.
Preferred Experience
- Demonstrated ability to lead cloud modernization migration or platform transformation initiatives in a regulated or enterprise environment.
- Hands-on experience with Kubernetes platform operations GitOps workflows service mesh secrets management and policy-as-code practices.
- Experience applying FinOps principles to optimize cloud consumption cost visibility capacity planning and resource governance.
- Ability to influence architecture and technical strategy across cross-functional teams while balancing reliability security scalability and business delivery needs.
- Strong analytical and problem-solving skills with the ability to think strategically and provide innovative solutions to complex security challenges.
Additional Information :
AbbVie is an equal opportunity employer and is committed to operating with integrity driving innovation transforming lives and serving our community. Equal Opportunity Employer/Veterans/Disabled.
US & Puerto Rico only - to learn more visit & Puerto Rico applicants seeking a reasonable accommodation click here to learn more:
Yes
Employment Type :
Full-time
About Company
AbbVie is a global biopharmaceutical company focused on creating medicines and solutions that put impact first for patients, communities, and our world. We aim to address complex health issues and enhance people's lives through our core therapeutic areas: immunology, oncology, neuro ... View more