Site Reliability Engineering Lead (Local Development Centre)
Job Summary
Leidos is seeking an experienced Site Reliability Engineering (SRE) Lead to support the Local Development Centre (LDC) in Singapore. This role provides technical leadership for system reliability observability performance engineering operational readiness and service management while working closely with customer engineers in a co-development and capability-building environment.
The successful candidate will establish reliability engineering practices drive automation and operational excellence initiatives improve system performance and availability and contribute to long-term engineering capability development through mentoring knowledge transfer and hands-on collaboration.
Leidos is a global technology company delivering innovative solutions across aviation transportation defense government and critical infrastructure Singapore Leidos partners with customers industry and research organizations to solve complex operational challenges through advanced engineering and digital transformation.
- Lead reliability and operational excellence initiatives for complex technology platforms.
- Influence monitoring observability automation and service management practices.
- Mentor engineers and develop long-term customer capability.
- Work closely with customers in a collaborative co-development environment.
- Collaborate with engineering teams across Singapore and the United States.
- Lead reliability engineering monitoring observability and operational support initiatives.
- Provide technical leadership and mentoring to site reliability engineers within the LDC.
- Establish reliability standards operational processes service level objectives (SLOs) and engineering best practices.
- Drive co-development activities with customer engineers providing guidance coaching and technical oversight.
- Collaborate with software infrastructure DevSecOps systems integration cybersecurity and test teams to improve system reliability and performance.
- Support incident management root cause analysis problem management and continuous improvement activities.
- Lead performance analysis capacity planning availability management and operational readiness assessments.
- Support development of automation solutions to improve monitoring reliability recovery and operational efficiency.
- Participate in architecture reviews technical planning risk assessments and delivery activities.
- Support knowledge-transfer activities technical training workshops and on-the-job training programs.
- Ensure operational environments meet reliability availability performance and maintainability requirements.
- Champion automation observability and operational excellence practices across the engineering organization.
- Bachelors degree in Computer Science Information Technology Systems Engineering Computer Engineering or a related technical discipline.
- 7 years of experience in site reliability engineering platform engineering infrastructure engineering operations or related disciplines.
- Experience leading operational or reliability engineering teams on complex technology programs.
- Strong understanding of monitoring observability performance management incident response and service management practices.
- Experience supporting highly available scalable and resilient technology platforms.
- Experience working in Agile delivery environments and cross-functional engineering teams.
- Experience mentoring engineers and supporting knowledge transfer workforce development training technical capability development.
- Strong communication problem-solving and stakeholder management skills.
- Ability to work effectively in customer-facing and cross-functional environments.
- Experience leading co-development teams involving customers partners or contractor personnel.
- Experience with observability monitoring logging and operational analytics platforms.
- Experience leveraging AI-assisted engineering tools to improve operational efficiency troubleshooting and engineering productivity.
- Experience using Agile planning and collaboration platforms such as Jira Azure DevOps Rally Confluence GitLab or similar tools.
- Scrum Master (CSM PSM) or other Agile-related certifications are desirable.
- Experience working in regulated mission-critical safety-critical or highly available operational environments.
- Aviation transportation defense government or critical infrastructure domain experience is desirable.
- Proficiency in English and one or more regional languages is desirable.
- Candidates with existing authorization to work in Singapore.
If youre looking for comfort keep scrolling. At Leidos we outthink outbuild and outpace the status quo because the mission demands it. Were not hiring followers. Were recruiting the ones who disrupt provoke and refuse to fail. Step 10 is ancient history. Were already at step 30 and moving faster than anyone else dares.
For U.S. Positions: While subject to change based on business needs Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.
The Leidos pay range for this job level is a general guideline onlyand not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job education experience knowledge skills and abilities as well as internal equity alignment with market data applicable bargaining agreement (if any) or other law.
Required Experience:
IC
About Company
Leidos is an innovation company rapidly addressing the world's most vexing challenges in national security and health. Our 47,000 employees collaborate to create smarter technology solutions for customers in these critical markets.