Level

Thales

Site Reliability Engineer

AI in this role

Location: Noida, India

Thales is a global technology leader trusted by governments, institutions, and enterprises to tackle their most demanding challenges. From quantum applications and artificial intelligence to cybersecurity and 6G innovation, our solutions empower critical decisions rooted in human intelligence. Operating at the forefront of aerospace and space, cybersecurity and digital identity, we’re driven by a mission to build a future we can all trust.

Present in India since 1953, Thales is headquartered in Noida and has other operational offices and sites spread across Delhi, Gurugram, Bengaluru and Mumbai, among others. Over 2200 employees are working with Thales and its joint ventures in India. Since the beginning, Thales has been playing an essential role in India’s growth story by sharing its technologies and expertise in Defence, Aerospace and Cyber & Digital sectors. Thales has two engineering competence centres in India - one in Noida focused on Cyber & Digital business, while the one in Bengaluru focuses on hardware, software and systems engineering capabilities for both the civil and defence sectors, serving global needs. The Group has also established an MRO (Maintenance, Repair & Overhaul) facility in Gurugram to provide comprehensive avionics maintenance and repair services to Indian airlines and support the growth of the local aviation industry.

Missions and Responsibilities

Reliability & Operations

• Ensure high availability, performance, and scalability of Thales PAY Digital cloud services in line with customer SLAs/SLOs

• Respond to production incidents, lead troubleshooting efforts, and coordinate recovery actions

• Conduct post-incident analyses (RCA/postmortems) and implement long-term corrective and preventive actions

• Participate in on-call and incident escalation processes

Automation & Infrastructure

• Design, develop, and maintain Infrastructure as Code (IaC) and automation solutions (Terraform, GitLab CI/CD, scripting)

• Support and validate cloud deployments of Thales PAY products (AWS / GCP / Kubernetes)

Change & Release Management

• Provide technical expertise and risk assessment during Change Advisory Board (CAB) activities, especially for high-risk changes impacting customer SLAs

• Ensure proper change planning, validation, and rollback strategies

• Contribute to release readiness and operational acceptance

Observability & Performance

• Implement and maintain monitoring, alerting, and observability across platforms (metrics, logs, traces)

• Continuously improve system monitoring, dashboards, and alert quality using tools such as Datadog and Splunk

• Perform capacity planning, performance tuning, and continuous technological watch

Collaboration & Documentation

• Work closely with Product Owners and development squads to anticipate operational needs

• Provide technical input for new services and service evolutions

Skills requirements (Domain/technical/soft)

Technical Skills

• Experience in systems, networking, and security operations

• Hands-on experience with Kubernetes, AWS, and/or GCP in production environments

• Strong experience with CI/CD pipelines and Infrastructure as Code (GitLab CI, Terraform)

• Proficiency in Linux, TCP/IP, HTTP/HTTPS, and distributed systems

• Strong knowledge of observability and monitoring tools (Datadog, Splunk, logs/metrics/traces)

• Solid scripting skills (Shell and/or Python)

Methodologies & Practices

• Knowledge of SRE principles (SLIs, SLOs, error budgets, automation, incident management)

• Familiarity with Agile and DevOps ways of working

• Experience in service delivery, operational readiness, and production support

Soft Skills

• Strong analytical and troubleshooting skills

• Ability to diagnose and resolve complex, high pressure production issues

• Clear communication with both technical and non-technical stakeholders

• Ownership mindset and customer-oriented approach


At Thales, we’re committed to fostering a workplace where respect, trust, collaboration, and passion drive everything we do. Here, you’ll feel empowered to bring your best self, thrive in a supportive culture, and love the work you do. Join us, and be part of a team reimagining technology to create solutions that truly make a difference – for a safer, greener, and more inclusive world.

How we score this

Site Reliability Engineer at Thales scores 38 out of 100 on AI centrality, which makes it AI Level 1 of 4 (Little AI) on this board. The level measures how much of the work is AI, not seniority.

Classification

AI Level 1. The work itself involves no AI, or AI only appears as scenery, such as a company tagline.

  1. AI Level 480 to 100
  2. AI Level 360 to 79
  3. AI Level 240 to 59
  4. AI Level 10 to 39

Bands come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.

Get new AI jobs by email

One email a week with the new AI jobs, each rated AI Level 1 to 4 for how much AI is in the work. No recruiter spam, unsubscribe in one click.

Free. One email a week. Unsubscribe in one click.

Similar roles

Software Engineering roles rated AI Level 1 at other companies.

What kind of AI work fits you?

Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.

Find my next step

More jobs at Thales

Related searches

Same AI level