Senior Director of Production Engineering-Observability & Telemetry Platforms
Zscaler · San Jose, California, USA
About this role
**Senior Director of Production Engineering - Observability & Telemetry Platforms** **About Zscaler** Zscaler (NASDAQ: ZS) accelerates digital transformation through the Zero Trust Exchange™️ platform, protecting thousands of customers from cyberattacks and data loss. Distributed across 160+ public exchanges globally, it's the world's largest in-line cloud security platform. **The Role** We're seeking an exceptional engineering executive to lead our distributed Telemetry Pipelines and enterprise-wide Observability Platform. This San Jose, CA hybrid role (3 days/week) reports to executive engineering leadership with end-to-end strategic, technical, and operational ownership. Our cloud platform processes hundreds of billions of transactions and petabytes of telemetry data daily. **What You'll Do** • Define long-term roadmap and architectural strategy for the global Observability Platform • Architect and scale high-throughput, low-latency telemetry ingestion pipelines handling petabyte-scale data • Optimize platform performance and resource utilization while managing aggressive growth • Partner with SRE, Product, and incident management teams on SLOs and alerting • Lead and mentor world-class engineering teams across multiple global sites • Act as trusted partner to Product, Security, and Support leadership **Who You Are** • Empathetic leader who develops high-performing, diverse engineering teams • Strategic thinker with proven track record managing large organizations • Culture champion fostering psychological safety and engineering excellence • Operational perfectionist focused on automation and resilient architectures • Collaborative partner bridging technical details and business objectives **Minimum Qualifications** • 7+ years progressive engineering leadership managing 30+ engineers in hyper-scale SaaS/cloud environments • Deep expertise in observability systems: OpenTelemetry, Kafka, Flink, Vector, VictoriaMetrics, Grafana, Prometheus, ClickHouse, OpenSearch • Strong foundation in cloud-native paradigms, Kubernetes, microservices, and SRE methodologies • Experience leveraging AI/ML tools to optimize production engineering workflows • Proven ability to influence executives and drive cross-functional alignment **Preferred Qualifications** • Compliance, data retention, PII masking, and data governance expertise • Open-source contributions in observability (OpenTelemetry, Prometheus) • Chaos engineering and automated disaster recovery experience **Compensation & Benefits** Base Pay Range: $231,000 – $330,000 USD Comprehensive benefits include health plans, time off, parental leave, retirement options, education reimbursement, and in-office perks. **Location:** San Jose, CA (Hybrid - 3 days/week on-site)
Listing freshness
CronJobs last confirmed this listing 4h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.