# Principal SRE Engineer at Entain

AI Level 1, AI centrality 0 out of 100. Remote (Hyderabad, Telangana, India).

## Details

- Company: [Entain](https://jobsbylevel.com/companies/entain)
- AI level: AI Level 1 (score 0 out of 100)
- Location: Remote (Hyderabad, Telangana, India)
- Posted: October 7, 2026
- Apply: https://jobsbylevel.com/go/e0f390d5-2b96-4df1-84c9-ea8545f0679d

## Description

There’s nothing like a career at Entain India. We’re the technology powerhouse helping Entain – a world leader in sports and gaming – create products and experiences used by millions of customers across the globe. Whether you’re an engineer, product owner or customer care professional, unlock a world of new possibilities – for you and the future of sports and gaming. We are looking for a Principal Site Reliability Engineer(SRE) to set the technical direction for the reliability, performance, and scalability of our hybrid infrastructure. As the most senior technical authority in the SRE function, you will define architecture and engineering standards across cloud and on-premise environments, promote the use of modern SRE practices. This is a hands-on, technical role (~90%) with influence and technical mentorship responsibilities (~10%). You will operate as a force multiplier, raising the technical bar for the entire engineering organisation rather than managing a team directly. Technical Responsibilities (~90%) Architect scalable, secure, and cost-efficient cloud and hybrid solutions, and establish the patterns and reference implementations other teams build on. Define the observability strategy across the organisation: monitoring, logging, alerting, SLIs/SLOs and error budgets using Datadog, Elastic, OpenTelemetry, and similar tooling. Set standards for high availability, disaster recovery, and backup across hybrid environments, and validate them through failure testing. Partner with development, security, and platform teams to shape deployment pipelines (CI/CD) and GitOps workflows at scale. Establish and maintain organisation-wide technical documentation, runbooks, architectural decision records, and operational standards. Identify systemic reliability bottlenecks, lead complex root-cause analysis, and reduce operational load through automation and platform improvements. Troubleshoot and resolve the most complex, production issues, and lead major incident response. Provide deep 3rd line technical support and participate in the on-call rotation as an escalation point. Technical Leadership & Influence (~10%) Be a technical mentor for SREs and engineers across teams, raising the engineering bar through coaching and knowledge-sharing. Lead structured up skilling and enablement across Windows, Linux, AWS, and Kubernetes. Influence and align Product, Engineering, Security, Platform, and Operations teams around reliability goals and technical direction. Shape the SRE roadmap together with SRE Leadership and SRE Guild and lead prioritisation of main reliability and platform programmes. Communicate technical strategy and trade-offs to both engineering teams and senior partners, and lead accountability for reliability outcomes. Champion a culture of operational excellence, ownership, and continuous improvement across the organisation. Experience: 10+ years of experience in SRE, DevOps, Platform Engineering, or Cloud Engineering, with experience operating at a senior or principal level. Deep, hands-on expertise with On-prem and large-scale AWS environments. Expert-level experience running Kubernetes in production (EKS, EKS Anywhere, EKS Hybrid preferred), including group design, scaling, and hardening. Background leading migrations of IIS/.NET and Linux-based applications to cloud and Kubernetes. Expert-level Infrastructure-as-Code skills, Terraform, including module design and standards for large teams. Knowledge of GitOps (ArgoCD), automated deployments, and configuration management. Across Windows, Linux, networking, and distributed systems. Experience with observability platforms (Datadog, Prometheus, OpenTelemetry) and defining SLI/SLO/error-budget practices. Experience with security best practices in cloud and hybrid environments. Experience operating in 24/7 high-availability, mission-critical environments. Excellent communication and cross-team collaboration skills, with the ability to influence technical decisions without

The description is cut here. Read the full offer: https://jobsbylevel.com/jobs/principal-sre-engineer-at-entain-8f6eb1

Source: https://jobsbylevel.com/jobs/principal-sre-engineer-at-entain-8f6eb1

## Cite this page

Level. https://jobsbylevel.com/jobs/principal-sre-engineer-at-entain-8f6eb1.

Get job alerts: https://jobsbylevel.com/newsletter
