Observability Build SME
Job Summary
Role Summary
We are looking for an Observability Engineer to design implement and automate enterprise monitoring and observability solutions. The role focuses on improving platform reliability enabling proactive monitoring and driving observability adoption across application infrastructure and cloud environments.
Key Responsibilities
- Build and maintain observability solutions using Splunk Prometheus OpenTelemetry Tempo and Cribl.
- Develop dashboards alerts KPIs and service health monitoring.
- Drive automation initiatives using Python Ansible APIs and CI/CD pipelines.
- Support instrumentation of applications infrastructure and distributed tracing.
- Collaborate with application infrastructure and platform teams to define monitoring requirements.
- Troubleshoot incidents perform root cause analysis and improve monitoring effectiveness.
- Promote observability standards best practices and self-service capabilities.
Required Skills
- Hands-on experience with Splunk Prometheus OpenTelemetry Tempo and Cribl.
- Good knowledge of Python Ansible GitLab and CI/CD practices.
- Hands-on experience on Cloud services AWS & Azure.
- Experience with monitoring alerting dashboards logging metrics and tracing.
- Strong automation mindset and problem-solving skills.
- Excellent stakeholder management and communication skills.
Good to Have
- Nagios Zabbix
- Splunk ITSI
- Kubernetes and cloud observability
- ServiceNow integrations
- SRE practices
Experience
- 4-8 years in Observability Monitoring SRE DevOps or Platform Engineering roles.
Key Traits: Automation-focused customer-oriented proactive and capable of working effectively with both technical and business stakeholders.
#LI-7013About Company
NXP is a global semiconductor company creating solutions that enable secure connections for a smarter world.