Engineering Manager – SRE
Whitefield, India · Full-time
- Posted 2d ago
- From Zluri’s careers page
- Location
- Whitefield, India
- Type
- Full-time
- Level
- Senior
- Experience
- 8+ years
- Department
- Engineering
Opens the listing on zluri.keka.com
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
About the Role
We are looking for an Engineering Manager – SRE (Site Reliability Engineering) to lead and own the reliability, scalability, and operational excellence of our platform. You will build and grow a high-performing SRE team while staying technically hands-on across AWS, Kubernetes, databases, observability, automation, and incident management. You will define and drive SLOs, error budgets, on-call practices, and reliability standards across critical services.
You will partner closely with Data Engineering, Backend, Product, and Security teams to improve platform resilience, automate operational workflows, optimize cloud costs, and maintain security and compliance standards.
This is a high-impact leadership role for someone who thrives in ambiguous, fast-paced environments and enjoys building both scalable systems and strong engineering teams.
Responsibilities:
- Build, lead and grow a team of SREs: hiring, mentoring, career growth and performance.
- Own overall infrastructure stability: define SLOs for our critical services and platform, and use error budgets to drive reliability priorities.
- Own incident management end-to-end: on-call, incident command, blameless postmortems and follow-through.
- Own production infrastructure on AWS (EKS, Glue, Lambda, S3) and the health of our datastores (MongoDB, PostgreSQL, Redshift, Redis).
- Build monitoring and alerting that tells us when customer data is late or wrong, not just when a service is down.
- Reduce toil through automation, infrastructure-as-code and safe deployment practices.
- Own cloud costs and partner with Security to keep SOC 2 / ISO 27001 controls operational.
- Work with Data Engineering, Backend and Product teams to make reliability trade-offs explicit.
Requirements:
Must have:
- Must have 8+ years of experience in SRE, DevOps or infrastructure engineering, with 2+ years managing engineers.
- Deep hands-on experience with AWS and Kubernetes in production.
- Should have operated a high-volume NoSQL OR relational database in production (replication, failover, query and index tuning).
- Should have built on-call, incident management and SLO practices from the ground up.
- Strong experience with infrastructure-as-code (Terraform or similar) and CI/CD pipelines.
- Prior experience with FinOps / cloud cost optimization.
- Should be able to handle ambiguous problems well.
- Understanding of data security and privacy principles.
Good to Have:
- Experience with observability tools like Datadog, Sentry or Prometheus.
- Experience in a B2B SaaS company with enterprise customers and SLAs
- Strong programming skills in Python or Go for automation.
Skills they ask for
Pick one to see other roles that ask for it.
About Zluri
More roles at Zluri
See all 8- Customer Success EngineerBengaluru · SeniorCustomer Service · SeniorBengaluru, India2d
- Senior Software Engineer (Backend)Bengaluru · SeniorEngineering · SeniorBengaluru, India1w
- Senior Data EngineerBengaluru · SeniorEngineering · SeniorBengaluru, India1w
- Senior Product ManagerWhitefield · SeniorProduct Management · SeniorWhitefield, India1w
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.