Site Reliability Engineer III
AI in this role
Design and operate secure, highly available Azure cloud platforms and infrastructure using Site Reliability Engineering best practices.
Are you passionate about building reliable, secure and scalable cloud platforms?
Would you like to lead Azure reliability initiatives and help engineering teams adopt Site Reliability Engineering best practices?
About the Business
LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Insurance vertical, we provide customers with solutions and decision tools that combine public and industry specific content with advanced technology and analytics to assist them in evaluating and predicting risk and enhancing operational efficiency. Our insurance risk solutions help drive better data-driven decisions across the insurance policy lifecycle, all while reducing risk. You can learn more about LexisNexis Risk at https://risk.lexisnexis.com/insurance.
About the Role
As a Site Reliability Engineer, you will design, operate and continuously improve secure, highly available Azure platforms that support large-scale production environments. You will provide technical leadership for reliability initiatives, develop reusable automation and partner with engineering, security, architecture and product teams to improve resilience and operational performance.
Responsibilities
- Design, implement and maintain highly available Azure infrastructure, including Azure Kubernetes Service, Virtual Machines, Functions, App Services, Storage, Networking, Key Vault and Azure Monitor.
- Define and maintain Service Level Indicators, Service Level Objectives and Error Budgets, driving continuous improvements in platform reliability and availability.
- Lead root cause analysis and post-incident reviews, developing resilience patterns such as auto-scaling, self-healing, disaster recovery, multi-region failover and high-availability architectures.
- Design and maintain Infrastructure as Code using Terraform, build reusable platform components and automate manual operational processes using Python, PowerShell, Go or Bash.
- Operate and optimise AKS clusters, establish container security standards, implement GitOps practices using tools such as ArgoCD and manage upgrades, capacity and workloads.
- Develop monitoring, logging, alerting and distributed-tracing strategies, maintaining dashboards through Grafana and Azure Monitor to identify service degradation proactively.
- Build and support reliable, repeatable and auditable CI/CD pipelines using GitHub Actions and Azure DevOps, including blue/green and canary deployment strategies.
- Provide technical leadership, mentor SRE I and SRE II engineers, promote SRE practices across teams and participate in on-call rotations and major incident management.
Requirements
- Experience in Site Reliability Engineering, Cloud Engineering, DevOps or Platform Engineering, including supporting large-scale production environments.
- Expert knowledge of Microsoft Azure and strong experience with Kubernetes, preferably Azure Kubernetes Service.
- Advanced Terraform skills and experience designing reusable Infrastructure as Code components.
- Strong Git and GitOps practices, with experience using CI/CD platforms such as GitHub Actions and Azure DevOps.
- Strong Linux administration, scripting and programming skills using languages such as Python, PowerShell, Go or Bash.
- Experience with monitoring and observability platforms, distributed tracing, application performance monitoring and proactive alerting.
- Skills in incident management, root cause analysis, capacity planning, performance optimisation, disaster recovery testing and production operations.
- Technical leadership, strategic thinking, problem solving, effective communication, collaboration, customer focus and an automation-first approach. Azure, Kubernetes or Terraform certifications are welcomed.
Learn more about the LexisNexis Risk team and how we work
here
Primary Location Base Pay Range: Ireland - Dublin (Rockfield Central) €51,400 - €85,700.
We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits. Click here to access benefits specific to your location.
We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact 1-855-833-5120.
Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams here.
Please read our Candidate Privacy Policy.
We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.
USA Job Seekers:
How we rate this
Site Reliability Engineer III at RELX rates 10 out of 100 for how much of the daily work is AI. That makes it Little AI (AI Level 1 of 4). The level is about AI in the job, not seniority.
Little AI. AI is not part of the work.
- ●●●● Builds AI80 to 100
- ●●●○ Works on AI60 to 79
- ●●○○ Uses AI40 to 59
- ●○○○ Little AI0 to 39
Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.
Prepare for this job
A free preview built only from this posting: what it asks for, what you could be asked in an interview, and how to adjust your resume.
Skills and AI tools this role asks for
Questions you could be asked
- Tell me about a project where site reliability engineering was part of your work. What did you do?
- Tell me about a project where kubernetes was part of your work. What did you do?
- Tell me about a project where infrastructure as code was part of your work. What did you do?
- Tell me about a project where ci cd was part of your work. What did you do?
- Tell me about a project where cloud platforms was part of your work. What did you do?
Adapt your resume
- List these exact terms on your resume: Site Reliability Engineering, Kubernetes, Infrastructure As Code, Ci Cd, and Cloud Platforms. An applicant tracking system matches the wording, not the idea.
- Attach one line of real, concrete experience to at least one of them — a tool named with nothing behind it rarely survives a human read.
Want an expert to read your CV for this job?
Free. Send your CV and the role you want next. We reply by email within 2 to 4 business days.
Get a free CV reviewGet new AI jobs by email
One email a week with the new AI jobs, each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Software Engineering roles that involve little AI, at other companies.
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next step