Site Reliability Engineer
AI in this role
Job Description:
Job Title: Site Reliability Engineer
Corporate Title: Assistant Vice President
Location: Bangalore, India
Role Description
We are looking for Site Reliability Engineer candidate with below requirement. This role is combination of Production support + SRE + Devops. So majorly looking for GCP experience and Kubernetes to support the design, deployment, automation, and operational excellence of enterprise-grade cloud applications.
What we’ll offer you
As part of our flexible scheme, here are just some of the benefits that you’ll enjoy,
Best in class leave policy.
Gender neutral parental leaves
100% reimbursement under childcare assistance benefit (gender neutral)
Sponsorship for Industry relevant certifications and education
Employee Assistance Program for you and your family members
Comprehensive Hospitalization Insurance for you and your dependents
Accident and Term life Insurance
Complementary Health screening for 35 yrs. and above
Your key responsibilities
System Reliability: Ensure the reliability, availability, and performance of production systems by implementing best practices in monitoring, alerting, and incident response.
System Maintenance: Understand thoroughly the end-to-end application support process and escalation procedures, become fully conversant with all support tools. Maintain an end-to-end view of the application and infrastructure landscape.
Automation: Develop and maintain automation tools and scripts to streamline deployment, scaling, and operational tasks.
Incident Management: Act as a primary responder to system outages and incidents, ensuring rapid resolution and thorough post-mortem analysis to prevent recurrence.
Monitoring & Alerting: Design and implement robust monitoring and alerting systems to proactively identify and address potential issues.
Performance Optimization: Identify and resolve performance bottlenecks across the stack, from application code to infrastructure.
Collaboration: Work closely with development teams and other stakeholders to ensure that new features and services are designed with reliability and scalability in mind.
Documentation: Maintain comprehensive documentation of systems, processes, and procedures to ensure knowledge sharing and continuity.
Continuous Improvement: Continuously evaluate and improve our infrastructure, tools, and processes to enhance system reliability and operational efficiency.
Design, implement, and manage CI/CD pipelines using GitHub Actions.
Deploy and operate applications on Google Kubernetes Engine (GKE).
Develop and maintain Helm charts for complex application deployments.
Manage Kubernetes infrastructure including node management, auto-scaling, configuration management, and secrets management.
Configure and support service networking components such as gateways, virtual services, and service mesh technologies (Anthos Service Mesh preferred).
Your skills and experience
Proficiency in Infrastructure as Code - Terraform (must)
Proficiency in cloud platforms such Google Cloud (preferred), Openshift Cloud
Usage of enterprise Security Management solutions including GCP Secret Manager.
Expertise in Kubernetes (GKE) administration and operations
Experience in CI/CD tools
GitHub Actions – CI/CD experience is must
Experience with Docker/Kubernetes (creating images, deployments)
Experience into developing Helm Charts ( templates, hooks, packaging)
Exposure to delivering good quality code within enterprise scale development
Working knowledge of environment monitoring tools such as GCO, Prometheus, Grafana
Strong experience in software development processes, models, lifecycles and methodologies.
Expert hands-on experience with service-mesh technology such as Istio or Anthos Service Mesh
Experience in software development and scripting in at least one language (Java, JavaScript, Python, Go, Bash)
Proven ability to leverage AI tools to enhance productivity, optimise workflows to solve business problems, while applying critical judgment to ensure responsible and ethical use of data and AI outputs.
How we’ll support you
Training and development to help you excel in your career.
Coaching and support from experts in your team.
A culture of continuous learning to aid progression.
A range of flexible benefits that you can tailor to suit your needs.
About us and our teams
Please visit our company website for further information:
https://www.db.com/company/company.html
We strive for a culture in which we are empowered to excel together every day. This includes acting responsibly, thinking commercially, taking initiative and working collaboratively.
Together we share and celebrate the successes of our people. Together we are Deutsche Bank Group.
We welcome applications from all people and promote a positive, fair and inclusive work environment.
How we score this
Site Reliability Engineer at Deutsche Bank scores 30 out of 100 on AI centrality, which makes it AI Level 1 of 4 (Little AI) on this board. The level measures how much of the work is AI, not seniority.
AI Level 1. The work itself involves no AI, or AI only appears as scenery, such as a company tagline.
- AI Level 480 to 100
- AI Level 360 to 79
- AI Level 240 to 59
- AI Level 10 to 39
Bands come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.
Get new AI jobs by email
One email a week with the new AI jobs, each rated AI Level 1 to 4 for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Software Engineering roles rated AI Level 1 at other companies.
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next step