Level

Verve Group

Senior SRE/DevOps Engineer

AI in this role

Who We Are

Verve has created a more efficient and privacy-focused way to buy and monetize advertising. Verve is an ecosystem of demand and supply technologies fusing data, media, and technology together to deliver results and growth to both advertisers and publishers–no matter the screen or location, no matter who, what, or where a customer is. With 30 offices across the globe and with an eye on servicing forward-thinking advertising customers, Verve’s solutions are trusted by more than 90 of the United States’ top 100 advertisers, 4,000 publishers globally, and the world’s top demand-side platforms. Learn more at verve.com.

Senior SRE/Devops Engineer

 

Who We Are

Verve Group has created a more efficient and privacy-focused way to buy and monetize advertising. Verve Group is an ecosystem of demand and supply technologies fusing data, media, and technology together to deliver results and growth to both advertisers and publishers–no matter the screen or location, no matter who, what, or where a customer is. With 13 offices across the globe and with an eye on servicing forward-thinking advertising customers, Verve Group’s solutions are trusted by more than 90 of the United States’ top 100 advertisers, 4,000 publishers globally, and the world’s top demand-side platforms. Learn more at www.verve.com.

About This Job

Verve's Systems Engineering team keeps a high-throughput, real-time ad-serving platform fast and available, running on GCP and Kubernetes (GKE) and serving traffic around the clock. We're looking for an engineer who treats reliability as a software problem, measuring it, engineering for it, and automating away whatever gets in the way. The role combines hands-on ownership of production systems with building tools and services that other engineers rely on every day. You'll define what "reliable" means for our services, build the observability to prove it, and write the code that keeps systems healthy at scale. When something is manual, fragile, or slow, you're the person who writes the code to fix it. You'll have real ownership: shaping how our platform evolves, making architectural calls on reliability and scalability, and working closely with backend and data teams to design services that hold up under real-world load.
What You'll Do
  • Define, measure, and uphold SLOs and error budgets for critical services, and use them to guide engineering priorities
  • Design and operate highly available, fault-tolerant infrastructure on GCP and Kubernetes for high-throughput, latency-sensitive workloads
  • Build observability across metrics, logs, tracing, and alerting that surfaces real user impact and cuts noise
  • Lead incident response, run blameless postmortems, and drive the fixes that stop problems from recurring
  • Reduce toil by writing tools, services, and automation (in Go, Python, or similar) that replace manual operational work
  • Own capacity planning, performance tuning, and cost efficiency as traffic grows
  • Partner with product, backend, and data teams on production readiness, resilient architecture, and safe release practices (progressive rollouts, automated rollback)
  • Participate in on-call and continuously improve its sustainability
What You Will Bring
  • Hands-on experience keeping large-scale, cloud-native production systems reliable
  • Strong experience with Google Cloud Platform and Kubernetes (GKE) in production, including how they behave under load and how they fail
  • Practical experience with SLIs, SLOs, error budgets, and incident management
  • Solid software development skills in Go / Python / Rust, or similar, with good coding practices (testing, code review, maintainability) and a track record of shipping tools or services others depend on
  • Infrastructure-as-code experience with Terraform and cloud APIs
  • Experience with observability tooling such as Grafana, VictoriaLogs, and PagerDuty
  • Experience with CI/CD and GitOps tooling such as Argo CD and GitHub Actions
  • Experience operating databases in production
  • Good understanding of networking, event messaging, and distributed systems failure modes
  • Strong troubleshooting skills across the stack, from Linux and networking up to application behavior
  • Full business proficiency in English
  • Bachelor's degree in Computer Science or equivalent practical experience

What We Offer

  • Just a few of the benefits waiting for you at Verve:

    • Keep healthy through our Apollo Health Check-ups, and stay covered with our Medical Insurance for you and your family

    • Pick what matters most to you in our INR 4100/month Personalized Benefits Platform: Travel, entertainment, food, fitness, and healthcare

    • Have us cover your office lunches 3 times a week through Zomato, and enjoy office sweets, snacks, and beverages

    • Boost your professional knowledge with our annual training budget & internal webinars, and level up your language skills through our German/English classes

    • Recharge with 19 paid vacation days + 3 Wellness days throughout the year, in addition to the public holidays

    • Strengthen team connections while exploring new cultures through our monthly Work&Travel raffle, offering a chance to work from one of our global offices (after 2 years of employment)

    • … and even more reasons to join us!

Verve provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. 

If you are a California resident subject to the California Consumer Privacy Act, click here to understand how Verve processes your personal information and how you can exercise your rights.
If you are located in the EU or UK visit our privacy policy to understand how Verve processes your personal information and how you can exercise your rights.

How we rate this

Senior SRE/DevOps Engineer at Verve Group rates 34 out of 100 for how much of the daily work is AI. That makes it Little AI (AI Level 1 of 4). The level is about AI in the job, not seniority.

Classification

Little AI. AI is not part of the work.

  1. ●●●● Builds AI80 to 100
  2. ●●●○ Works on AI60 to 79
  3. ●●○○ Uses AI40 to 59
  4. ●○○○ Little AI0 to 39

Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.

Get new AI jobs by email

One email a week with the new AI jobs, each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.

Free. One email a week. Unsubscribe in one click.

Similar roles

Software Engineering roles that involve little AI, at other companies.

What kind of AI work fits you?

Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.

Find my next step

More jobs at Verve Group

Related searches

Same AI level