MyInternships.in
Zeta Global logo — Zeta Global Lead Site Reliability Engineer at Zeta Global
Zeta Global

Lead Site Reliability Engineer

Job · Full-timeIn OfficeBengaluru3-5 years1 opening

About this role

Zeta Global is hiring for Lead Site Reliability Engineer in Bengaluru. This opening was published by Zeta Global on their official careers board (Greenhouse) on 15 July 2026 and was confirmed live on 14 September 2026. Job details • Company: Zeta Global • Role: Lead Site Reliability Engineer •…

Similar jobs hiring now

Not quite right? These jobs in Bengaluru match the same skills — apply to a few to improve your chances.

Skills you'll use

PythonAWSKubernetesLinuxGoTerraformCI/CDNetworkingShell Scripting

What you'll do

  • Implement and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets to drive reliability efforts.
  • Develop systems that are resilient to failures and ensure 99.9%+ uptime for critical services.
  • Lead incident response and post-incident reviews (blameless postmortems), ensuring robust root cause analysis and continuous improvement of systems.
  • Automate incident detection and response using automated runbooks or predefined workflows.
  • Write software as needed to support reliability or efficiency needs.
  • Design and implement full observability across systems using modern tools like Open Telemetry for tracing, metrics, and logging.
  • Use capacity planning, forecasting, and performance testing to ensure that the systems scale effectively as the user base and load grow.
  • Collaborate with development and operations teams on building reliable, scalable, and high-performance services.
  • Ensure best practices are followed across infrastructure design, deployment, and maintenance using tools like AWS, Kubernetes, EKS, Fargate, etc.
  • Champion Infrastructure as Code (IaC) to provision, manage, and scale infrastructure using tools like Terraform, Pulumi, or similar.

Who can apply

  • Experience & Qualifications:
  • 3-5 years of experience as an SRE, working in cloud-based environments and on-prem environments.
  • Deep understanding of Linux systems, networking, and systems administration.
  • Experience with cloud platforms like AWS, with a strong understanding of Kubernetes and container orchestration tools.
  • Hands-on experience with observability tools such as Honeycomb, Grafana, Prometheus, Thanos, ELK (Elastic Stack), or Loki.
  • Strong skills in at least one programming language (Python, Go) to write production level code.
  • Strong skills in shell scripting using bash or similar. • Experience with OpenTelemetry or other distributed tracing systems, including tracing, metrics, and logs integration.
  • Experience with Chaos Engineering methodologies and tools (Chaos Mesh, chaos monkey, AWS Fault Injection Simulator, etc.

About Zeta Global

About Zeta Global • Zeta Global unifies identity, data and AI in one marketing cloud, so enterprise brands can acquire, grow and retain customers across every channel. • The intelligence layer for agentic marketing and the enterprise, guiding every decision, action, and outcome. • Zeta Live…
Visit Zeta Global’s official website ↗

Searching for “Zeta Global jobs” or “jobs at Zeta Global”? You’re in the right place.

Sponsored — Deals for Professionals

Sponsored

Similar Jobs

Hand-picked roles that match this job's skills

Sponsored — Deals for Professionals

Sponsored

Similar Jobs Based on Your Skills

Recently Posted Jobs in Bengaluru

Fresh jobs posted in Bengaluru — apply early

Ready to apply for Lead Site Reliability Engineer?

Free to apply · takes under 2 minutes · Zeta Global reviews on a rolling basis.

Apply Now