# Site Reliability Engineer, AVP at Deutsche Bank

Deutsche Bank is hiring a Site Reliability Engineer, AVP. Level rates it Little AI ●○○○; you can [apply on Level](https://jobsbylevel.com/go/9e0c7585-2414-49f6-9917-6b376d4b059c).

AI Level 1, AI centrality 0 out of 100. (do not use) Bangalore, Velankani Tech Park.

## Details

- Company: [Deutsche Bank](https://jobsbylevel.com/companies/deutsche-bank)
- AI level: AI Level 1 (score 0 out of 100)
- Location: (do not use) Bangalore, Velankani Tech Park
- Posted: October 9, 2026
- Apply: https://jobsbylevel.com/go/9e0c7585-2414-49f6-9917-6b376d4b059c

## Description

Job Description: Job Title: Site Reliability Engineer, AVP Location: Bangalore, India Corporate Title: AVP Role Description We are looking for a hands-on AVP Site Reliability Engineer who is passionate about Kubernetes platforms, deep troubleshooting, automation, and continuous improvement. You will diagnose and resolve complex production issues across application, platform, network, storage, and infrastructure layers, connecting system behavior, observability signals, and tenant impact to drive root-cause resolution. Our SREs are thoughtful and pragmatic engineers who balance doing things right with doing what is needed now. Success in this role requires technical depth, curiosity, sound judgement, clear communication, and the ability to independently drive durable solutions across complex domains while collaborating effectively with global engineering teams. About CaaS Private CaaS Private is an advanced on-premises Kubernetes platform built on Google Distributed Cloud (GDC). The platform hosts critical applications. It provides a resilient, observable, and secure container platform for workloads that require enterprise-grade reliability and operational discipline. What we’ll offer you As part of our flexible scheme, here are just some of the benefits that you’ll enjoy Best in class leave policy Gender neutral parental leaves 100% reimbursement under childcare assistance benefit (gender neutral) Sponsorship for Industry relevant certifications and education Employee Assistance Program for you and your family members Comprehensive Hospitalization Insurance for you and your dependents Accident and Term life Insurance Complementary Health screening for 35 yrs. and above Your key responsibilities Define, implement, and continuously improve service-level indicators, service-level objectives, alerting standards, and error budgets for the CaaS Private platform and its critical services. Build and maintain observability across metrics, logs, traces, alerts, and dashboards to provide clear insight into platform health, saturation, latency, capacity, and failure modes. Investigate and resolve complex production issues across Kubernetes, Linux, networking, storage, ingress, service mesh, node services, platform dependencies, and tenant workloads. Lead or coordinate incident response for platform-impacting events, ensuring timely mitigation, clear stakeholder communication, blameless post-incident reviews, and durable follow-up actions. Automate repetitive operational tasks, diagnostics, and remediation workflows to reduce toil, improve consistency, and accelerate recovery. Improve platform reliability, upgrade safety, resilience, capacity planning, performance, disaster recovery, and operational readiness. Improve runbooks, dashboards, alerts, support processes, and engineering standards based on recurring issues and operational learning. Partner with platform, network, security, storage, and application teams on release readiness, troubleshooting, change execution, documentation, and adoption of SRE practices. Contribute to reliability-focused platform enhancements and use system-level insights to improve both platform stability and tenant experience. Support the global GDC/Kubernetes platform in a 24x7 follow-the-sun model, including on-call, weekend, public-holiday rotation, and early Monday coverage as required. Your skills and experience A bachelor’s degree in a technical or engineering discipline, with 10–12 years of hands-on experience in Site Reliability Engineering, Production Engineering, DevOps or a closely related infrastructure role. Strong hands-on Kubernetes expertise, including cluster operations, upgrades, troubleshooting, networking, storage, security, Helm, Operators, and workload lifecycle management in bare-metal or private-cloud environments. Demonstrated SRE mindset and experience with SLI/SLO/SLA management, error budgets, incident reduction, capacity planning, performance optimization, resilience engineering,

The description is cut here. Read the full offer: https://jobsbylevel.com/jobs/site-reliability-engineer-avp-at-deutsche-bank-558ae0

Source: https://jobsbylevel.com/jobs/site-reliability-engineer-avp-at-deutsche-bank-558ae0

## Cite this page

Level. https://jobsbylevel.com/jobs/site-reliability-engineer-avp-at-deutsche-bank-558ae0.

Get job alerts: https://jobsbylevel.com/newsletter
