CronJobs

devops-sre jobs

Staff Engineer, Site Reliability

Babylist · United States

remotestaff$226,673–$271,991Posted Sep 3, 2026AWSTerraformKubernetesMySQLRedisRuby on RailsSidekiqCircleCI

Apply on the employer site

About this role

**How We Build** Babylist is rebuilding engineering culture around a simple belief: **AI changes everything**—how teams are structured, how decisions get made, and how fast ideas become working software. Engineers own problems end to end with short feedback loops and real stakeholder access. We ship, learn, and iterate quickly, and when something isn’t working, we throw it out and start over. AI tools are a natural part of our workflow (like an IDE or version control). We use AI to explore tradeoffs, pressure-test designs, and move from problem to solution in hours instead of days—so engineers can focus on the decisions that require human judgment. --- **Our Tech Stack** - Ruby on Rails - AWS - Sidekiq - MySQL - Redis --- **What the Role Is** Babylist’s **Platform team** is the foundation every engineering team builds on. As a **Staff SRE**, you’ll own the infrastructure and reliability practices that support **9M+ users** and the engineers who build for them. This is **not** a maintenance role. You’ll actively evolve how we build and operate **AWS infrastructure, CI systems, and developer tooling**, with cross-functional impact across Babylist Engineering. --- **Who You Are** - **Deep hands-on Terraform expertise** — you own IaC, not just contribute to it - **Proven AWS experience at scale** — EKS, RDS, cloud networking, DNS, CDNs, load balancers (and the gotchas) - **Experienced operating Kubernetes in production** — you’ve debugged the hard stuff - **Comfortable designing and improving CI/CD systems** — CircleCI, GitHub Actions, or similar; you care about developer velocity - **Strong observability instincts** — Datadog, Sentry, PagerDuty, Cronitor; alerting that’s actionable, not noisy - **Experienced with on-call and incident management** — you’ve run post-mortems and changed things afterward - **Comfortable supporting developers** across local, staging, and production — you’re a resource, not a gatekeeper - **You naturally reach for AI in your work** — teams use AI daily, and you stay curious about what’s next --- **How You Will Make an Impact** - **Infrastructure ownership** — manage and evolve AWS using Terraform; keep EKS clusters, databases, and core services current and performant - **CI/CD reliability** — own speed and reliability for the full Engineering org; every deploy starts here - **Developer support** — be the go-to person when environments break; unblock teams fast across local, staging, and production - **Monitoring & alerting standards** — establish and socialize best practices so the right people get paged for the right reasons - **Incident response** — lead or support incidents, drive post-incident reviews, and close the loop - **Platform strategy** — contribute to architectural decisions shaping Babylist’s infrastructure over the next several years --- **Why This Role** - Platform is a team every engineer depends on—your work has **outsized leverage** across the product org - The infrastructure is solid but actively evolving—you’re shaping what comes next - Staff-level role with real cross-team visibility—you influence how Babylist engineers build and ship - You’ll support systems used by millions of families at a high-stakes life moment --- **About Compensation** We use a market-based approach. The starting salary range for this role is: **$226,673 to $271,991 + target 20% annual bonus and competitive equity** Starting salary depends on location, experience, and qualifications, with increases ov

Listing freshness

CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord