ASAI job platform logo
  1. Home
  2. /Jobs
  3. /DevOps Engineer Jobs in Bengaluru
  4. /Lead Site Reliability Engineer
Zeta Global

Lead Site Reliability Engineer

Bengaluru · Mid Level

Older listing - lower visibility likelyVerified listingPosted 82d ago

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

What we know about this role

Hiring pulse

MEDIUM

Zeta Global is reviewing applications at a steady pace. Expect a standard response time as they evaluate the current pool.

Apply window

First 72 hours

Window passed - posted 82d ago

Early applicants get seen before the pile builds.

Not a repost

The first time we've seen this listing - it hasn't been closed and reopened.

Skills required

19 listed
Service Level ObjectivesCI/CDService LevelPerformance TestingChaos EngineeringChaos Monkey (Software)Fault InjectionCircuit Breakers+11 more
Service Level ObjectivesCI/CDService LevelPerformance TestingChaos Engineering

You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.

Job description

Lead Site Reliability Engineer

Key Responsibilities:

  • Implement and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets to drive reliability efforts.

  • Develop systems that are resilient to failures and ensure 99.9%+ uptime for critical services.

  • Lead incident response and post-incident reviews (blameless postmortems), ensuring robust root cause analysis and continuous improvement of systems.

  • Automate incident detection and response using automated runbooks or predefined workflows.

  • Write software as needed to support reliability or efficiency needs.

  • Design and implement full observability across systems using modern tools like Open Telemetry for tracing, metrics, and logging.

  • Use capacity planning, forecasting, and performance testing to ensure that the systems scale effectively as the user base and load grow.

  • Collaborate with development and operations teams on building reliable, scalable, and high-performance services.

  • Ensure best practices are followed across infrastructure design, deployment, and maintenance using tools like AWS, Kubernetes, EKS, Fargate, etc.

  • Champion Infrastructure as Code (IaC) to provision, manage, and scale infrastructure using tools like Terraform, Pulumi, or similar.

  • Get involved in chaos engineering initiatives.

  • Participate on our on-call rotation.

  • Drive advanced alerting and anomaly detection applied to metrics

Key Requirements:

Experience & Qualifications:

  • 3-5 years of experience as an SRE, working in cloud-based environments and on-prem environments.

  • Deep understanding of Linux systems, networking, and systems administration.

  • Experience with cloud platforms like AWS, with a strong understanding of Kubernetes and container orchestration tools.

  • Hands-on experience with observability tools such as Honeycomb, Grafana, Prometheus, Thanos, ELK (Elastic Stack), or Loki.

  • Strong skills in at least one programming language (Python, Go) to write production level code.

  • Strong skills in shell scripting using bash or similar. • Experience with OpenTelemetry or other distributed tracing systems, including tracing, metrics, and logs integration.

  • Experience with Chaos Engineering methodologies and tools (Chaos Mesh, chaos monkey, AWS Fault Injection Simulator, etc.

Skills & Knowledge:

  • Reliability-focused mindset with the ability to balance fast product iterations and system stability.

  • Solid understanding of SLOs, SLIs, and error budgets.

  • Hands-on knowledge of CI/CD pipelines and infrastructure automation.

  • Proven expertise in incident management, postmortems, and root cause analysis.

  • Knowledge of modern deployment strategies (e.g., blue-green deployments, canary releases) and resiliency patterns (circuit breakers, retry mechanisms, etc).

Preferred Qualifications:

  • Experience with distributed systems.

  • Experience with statistical analysis applied to metrics.

  • Familiarity with high-performance, low-latency systems.

  • Strong problem solving skills.

  • Experience as on-call engineer.

  • Experience writing production level code.

  • Hands-on experience running Chaos Engineering drills and initiatives.

Company Summary

Zeta Global is a data-powered marketing technology company with a heritage of innovation and industry leadership. Founded in 2007 by entrepreneur David A. Steinberg and John Sculley, former CEO of Apple Inc and Pepsi-Cola, the Company combines the industry’s 3rd largest proprietary data set (2.4B+ identities) with Artificial Intelligence to unlock consumer intent, personalize experiences and help our clients drive business growth.

Zeta Global is a leading AI-powered marketing technology company that enables enterprise brands to acquire, grow, and retain customers through intelligent, data-driven engagement. At the center of its innovation is the Zeta Marketing Platform (ZMP), which unifies customer data, identity, and advanced analytics to transform billions of data signals into actionable marketing intelligence and measurable business outcomes.

Publicly traded on the New York Stock Exchange (NYSE: ZETA), Zeta is redefining modern marketing with Athena by Zeta™, a superintelligent, conversational agent embedded in ZMP that personalizes the marketer’s workspace, surfaces platform generated insights through natural dialogue and recommends next-best actions to help brands accelerate and optimize their marketing outcomes.

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

More DevOps Engineer roles in Bengaluru

See all

Lead Developer - Java React AWS AI

Referrals Only · Bengaluru

Consultant_CloudOps

KPMG India · Bengaluru

Senior Backend Engineer - Cloud Infra & Platform engineer

Weekday · Bengaluru

IT Software Engineer, Infrastructure

Databricks · Bengaluru

Keep browsing

All open roles at Zeta GlobalAll DevOps Engineer jobs in Bengaluru

Two ways in

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

ASAI job platform logo

A job platform finally, balanced in your favour.

Jobs by CityJobs by CompanyGuidesHow We VerifyAboutPrivacy PolicyTerms of Service

Built in India 🇮🇳

© 2026 ASAI. All rights reserved.