See all roles

[Remote] Staff Site Reliability Engineer

Work from home Full-time role Hiring

Note: The job is a remote job and is open to candidates in USA. Zscaler accelerates digital transformation to ensure customers are agile, efficient, resilient, and secure. They are seeking a Staff Site Reliability Engineer to be a key member of the Zero Trust Exchange team, responsible for the reliability of large-scale cloud services and ensuring system performance and availability.

Responsibilities

  • Own the reliability of a large-scale cloud service (Linux/BSD, bare metal, Kubernetes, custom load balancing, SD-WAN) by partnering with Engineering and Network teams to define requirements early, conduct operability reviews, and contribute code/design docs for platform resilience
  • Develop and operate end-to-end observability (metrics/logs/traces, dashboards, alerting) and incident tooling to manage SLOs/error budgets, reduce noise, and improve system detection and diagnosis
  • Participate in an on-call rotation to lead full-cycle incident response; perform deep cross-stack troubleshooting (OS, networking, distributed systems, packet captures, core dumps) to drive permanent software fixes and codify learnings into runbooks and tests
  • Build and maintain everything-as-code for fleet and service lifecycle, driving provisioning, configuration, release automation, canary deployments, and complex rollout/rollback workflows
  • Continuously improve platform hygiene through consistent OS/app upgrades, dependency/vulnerability patching, capacity and performance tuning, and strict CI/CD validation prior to production rollouts

Skills

  • US Citizenship is required (due to the nature of assigned customers)
  • 5+ years industry experience in software engineering, infrastructure software, and/or platform engineering
  • Proficiency in at least one programming language (such as Python, Bash, or Go) with demonstrated ability to write production-quality code (testing, code reviews, CI, maintainable design, scripting for diagnostics)
  • Strong Linux/Unix systems fundamentals (process/memory, filesystems, networking stack basics, debugging/perf troubleshooting) and solid understanding of networking protocols and components (e.g., HTTP, DNS, TCP/IP, ICMP, OSI model, subnetting, and load balancing/traffic concepts)
  • Proven experience operating production services (including incident response, troubleshooting, reducing toil) and ability to participate in on-call rotations and support occasional after-hours or weekend deployments
  • Managing BSD in production, with a focus on driving systemic fixes through platform engineering
  • Proven expertise in operating Kubernetes at scale
  • Deep experience with the Prometheus/OpenTelemetry ecosystems, including instrumenting golden signals, defining SLOs, and performing alert tuning to ensure high-availability environments

Benefits

  • Various health plans
  • Time off plans for vacation and sick time
  • Parental leave options
  • Retirement options
  • Education reimbursement
  • In-office perks, and more!

Company Overview

  • Zscaler is a global cloud-based information security company that enables secure digital transformation for mobile and cloud. It was founded in 2008, and is headquartered in San Jose, California, USA, with a workforce of 5001-10000 employees. Its website is https://www.zscaler.com.
  • Apply To This Job

    You might like

    [Remote] Trading Operations Team Lead

    Work from home Full-time role

    [Remote] Senior Account Manager - Retail (Remote Atlanta)

    Work from home Full-time role

    [Remote] Senior Manager - Sports CRM Analytics

    Work from home Full-time role

    [Remote] Regional Deployment Project Manager - Oracle Health

    Work from home Full-time role

    [Remote] Customer Support Specialist, Payments

    Work from home Full-time role

    [Remote] Data Platform Administrator

    Work from home Full-time role

    [Remote] Senior Staff Engineer, Interactive Voice Response - AI/ML

    Work from home Full-time role

    [Remote] Principal Specialist, Program Control Analyst (Remote)

    Work from home Full-time role

    [Remote] Senior Engineer/Scientist - Statistician

    Work from home Full-time role

    [Remote] Business Systems Administrator ERP - Miller Electric Company

    Work from home Full-time role

    AI Product Manager w/ Pharma/BioTech/Med Dev exp Must

    Work from home Full-time role

    Experienced Part-Time Customer Service Representative – Remote Support Team at arenaflex

    Work from home Full-time role

    Amazon Remote work From Home Job - Part-time job

    Work from home Full-time role

    Product Marketing Director

    Work from home Full-time role

    Online Remote Customer Service Representative at Southwest Airlines (Online Remote jobs)

    Work from home Full-time role

    Experienced Full Stack Customer Service Representative – Remote Support for arenaflex

    Work from home Full-time role

    Senior Technical Product Manager

    Work from home Full-time role

    Fraud Waste and Abuse - Sr. Analyst

    Work from home Full-time role

    Residential & Airbnb Cleaners Wanted!

    Work from home Full-time role

    Experienced Customer Care Advisor – Remote Opportunity with arenaflex

    Work from home Full-time role