Zscaler
Production Engineer
About this role
Zscaler is seeking a Production Engineer to enhance the reliability and scalability of their cloud infrastructure, which processes over 200 billion transactions daily. This role focuses on adopting an 'automation-first' approach, improving observability, and driving incident management efforts within a collaborative team environment.
What you'll do
- Implement scalable infrastructure across AWS, GCP, and bare-metal environments
- Drive automation by coding in Python/Go to create self-healing systems
- Establish observability standards and define SLIs/SLOs
- Act as lead Incident Commander and develop response playbooks
- Conduct operability reviews with partner teams
What they're looking for
- Proficiency in programming (Python, Go, or C/C++)
- Experience managing large-scale production services
- Strong understanding of networking protocols and Linux systems
- Knowledge of ITIL frameworks for service maturity
- Familiarity with observability tools (Prometheus, Grafana)
- Experience in incident management and on-call rotations
Benefits
- Remote work options available
- Opportunities for personal and professional growth
- Dynamic team-focused culture
- Chance to impact cybersecurity in a meaningful way
- Work in a high-performing team environment
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Zscaler
Zscaler builds Zero Trust security platforms and managed detection and response (MDR) services that protect endpoints and cloud infrastructure while processing massive transaction volumes. The company is hiring threat response engineers for security operations, sales engineers to support enterprise customers, and production/reliability engineers to maintain their scalable cloud infrastructure.
- Website
- zscaler.com