CronJobs

devops-sre jobs

Senior Incident Commander

Box · Redwood City, CA, United States

hybridsenior$187,000–$187,000Posted Oct 2, 2026PythonGoKubernetesTerraformAWSGCPAzureLinux/Unix

Apply on the employer site

About this role

**Senior Incident Commander** **About Box** Box (NYSE:BOX) is the leader in Intelligent Content Management. Our platform enables organizations to fuel collaboration, manage the entire content lifecycle, secure critical content, and transform business workflows with enterprise AI. **The Role** Box is seeking a Senior Incident Commander with strong Site Reliability Engineering (SRE) expertise and proficient Python skills. You'll lead critical and blocker incidents to swift resolution while designing tools, automation, and processes for next-generation cloud operations. **Key Responsibilities** • Own live-site Critical and Blocker incidents from identification through mitigation and recovery • Triage problems, organize incident bridges, coordinate SMEs, and lead cross-functional teams • Improve incident platform tooling and automate repetitive response steps • Partner with SRE and engineering teams to implement secure, automated workflows • Lead daily change reviews and minimize change risk • Provide technical expertise in 24x7 global environments • Lead projects to enhance site resiliency, service manageability, and observability • Turn incident learnings into actionable engineering improvements • Mentor and uplift the team through exercises and runbook improvements **Required Qualifications** • 5+ years in SRE, production operations, or equivalent high-scale SaaS operations • Demonstrated Incident Commander/Technical Duty Officer experience • Strong SRE fundamentals: SLIs/SLOs, observability, blameless postmortems • Proficient Python for automation and tooling • Solid Linux/Unix troubleshooting and distributed systems knowledge • Networking literacy (DNS, TLS, load balancing, HTTP) • Cloud environment experience (GCP preferred) • Excellent written and verbal communication • Proven ability to coach and mentor others **Preferred Skills** • 24x7 NOC/GTOC operations experience • Hands-on with Prometheus, distributed tracing, synthetic monitoring, PagerDuty • Incident tooling improvement experience • Change management under pressure • Go, shell, Terraform, or CI/CD experience • Service catalog or dependency mapping experience **Culture & Location** Minimum 3 days per week in-office. Box values diversity and encourages applications from candidates of all backgrounds.

Listing freshness

CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord