Senior SRE Observability Engineer
Posted:
11 September 2026 (10 hours ago)
Application Deadline:
9 December 2026
Vacancies:
1 Vacancy
Job Summary
Greetings from Maneva!
Job Description
Senior SRE / Observability Engineer
Location: Hyderabad / Chennai / Bangalore
Experience: 8 - 15 Years
Notice Period: Immediate to 15 Days
Requirements:
- Provide SRE and Dynatrace subject matter expertise to mature observability across the Core Banking Vault ecosystem supporting a single pane of glass view across services dependencies and operational health.
- Work with upstream and downstream neighbours to strengthen end-to-end monitoring ensuring consistent telemetry ingestion dashboards alerts and actionable insight across critical service journeys.
- Expand monitoring capability across Istio Kafka synthetic monitoring and cloud cost observability with a focus on standardisation automation incident reduction and operational resilience.
- Establish and embed SLI/SLO practices including service level indicators service level objectives error budgets burn-rate alerting and evidence-based reliability improvements.
- Support ongoing training and knowledge transfer to build internal capability and improve consistent use of Dynatrace and observability practices across the platform. Essential Skills:-
- Strong experience in Site Reliability Engineering Observability Engineering Platform Engineering or DevOps within cloud-native and business-critical environments.
- Hands-on experience with Dynatrace including dashboarding alerting telemetry analysis synthetic monitoring and configuration management.
- Strong understanding of observability principles across metrics logs traces events service dependencies and user journeys.
- Experience applying observability configuration using self-service or automated platforms with good governance and standardisation.
- Strong working knowledge of Kubernetes platforms preferably GKE and service mesh technologies such as Istio.
- Experience using OpenTelemetry data including standardised ingestion reporting dashboards and alert generation.
- Strong working knowledge of Kafka observability including monitoring brokers topics partitions consumers lag throughput errors and resilience indicators.
- Experience defining and implementing SLI/SLO frameworks including burn-rate alerting error budgets and reliability dashboards.
Required Skills
- Strong experience refining alerts to reduce false positives and improve actionable alerting for genuine service issues.
- Good understanding of incident management problem management operational readiness and continuous improvement practices.
- Banking or core banking platform experience.
- Experience working with Thought Machine Vault.
- Experience supporting observability for large-scale distributed systems and regulated production platforms.
- Experience integrating observability data with cloud cost data to support FinOps insight and cost optimisation.
- Experience with GCP cost management FinOps practices cloud consumption reporting or cost-to-service mapping.
- Experience designing training material running enablement sessions and building internal knowledge across engineering teams.
- Knowledge of operational resilience frameworks service health reporting and evidence-based reliability improvement.
If you are excited to grab this opportunity please apply directly or share your CV atand