Sr. Staff Technical Program Manager
Zscaler
Let the right jobs find you
Get personalised suggestions from verified company career pages, matched to your role, location, level, and skills.
Overview
Position Type
Full Time
Experience
8+ years
Job Description
Role
We are looking for a Sr. Staff Technical Program Manager – Service Health to join our team. This role is for San Jose, CA (hybrid three days a week) reporting to the Senior Director of Site Reliability Engineering in the Cloud Operations department. Zscaler is seeking a Sr. Staff Technical Program Manager – Service Health to lead the central Service Health Program across the Zero Trust Exchange. You will define and govern service health metrics and operating cadences, align stakeholders, and leverage Critical User Journey (CUJ) and Lighthouse programs to drive durable operational improvements at global scale.
What you’ll do (Role Expectations)
- Own the company-wide Service Health Program from strategy through execution, defining charters, roadmaps, and success metrics across the Zero Trust Exchange
- Establish and govern service health practices by defining SLIs, SLOs, and error budgets to operationalize decision frameworks that balance feature velocity with reliability
- Drive Critical User Journey (CUJ) and Lighthouse signal adoption, converting operational alerts into explicit decisions and measurable engineering outcomes
- Build the service health source of truth by implementing continuous measurement dashboards across security, defects, incidents, capacity, availability, and latency
- Run continuous operating cadences and cross-organizational programs to standardize intake, prioritization, delivery tracking, and post-delivery verification at scale
Who You Are (Success Profile)
- You think at scale. You connect day-to-day execution to the overarching company mission, building durable processes and solutions designed for a high-growth global organization.
- You champion simplicity. You excel at distilling complex technical challenges, user needs, and operational concepts into clear, actionable plans that align teams.
- You are data-driven. You leverage metrics, telemetry, and evidence to establish truth, measure what matters, and guide informed decisions.
- You act like an owner. You operate with a strong bias for action and high integrity, moving seamlessly between strategic planning and hands-on execution.
- You are a high-trust collaborator. You foster alignment across engineering and leadership teams through transparent communication, constructive feedback, and shared accountability.
What We’re Looking for (Minimum Qualifications)
- Demonstrated curiosity and active exploration of AI tools, with a proven history of integrating new technologies to enhance daily workflows and augment problem-solving
- 8+ years in Technical Program Management, Production Engineering, or similar roles in large-scale, distributed systems
- Deep experience with SLIs/SLOs, error budgets, and service health practices with demonstrable outcomes improving availability, latency, and durability
- Proven leadership of cross-functional programs driving production service health at global scale across multi-product, multi-cloud, or multi-region environments
- Strong data fluency across instrumentation, metrics pipelines, synthetic monitoring, and executive dashboards to translate signal into actionable outcomes
- Track record of influencing engineers, managers, and senior leadership through clear tradeoff framing, written updates, and decision facilitation
What Will Make You Stand Out (Preferred Qualifications)
- Background partnering with Production Engineering on incident, problem, and post-incident programs, including RCA quality and preventive action systems
- Experience implementing operating cadences for service health reviews and OKRs across multiple product lines within networking or cybersecurity domains
- Familiarity with capacity planning, staged rollout safety guardrails, and reliability engineering practices such as SRE or ITIL