Senior DevOps Engineer (PST)
Scribe · Remote (PST)
About this role
About Scribe Scribe https://scribe.com is where exceptional people come to do the best work of their careers. More than 94% of the Fortune 500 use Scribe to own their specialized intelligence: the unique way their teams work, decide, and get things done. Our Specialized Intelligence platform automatically captures how work happens and turns it into a living asset that helps people and AI agents do their best work. We're growing fast. Since our founding in 2019, we've grown to 90,000+ customers and 7 million users across 600,000 businesses. Based in San Francisco, we've been named a LinkedIn Top Startup, are valued at over $1 billion, and are backed by leading investors. Join us in our mission to transform how people do work. About the role As our Senior DevOps Engineer, you'll own the infrastructure every engineering team at Scribe builds on top of. You'll set technical direction for reliability, performance, and deployment across AWS and Kubernetes, with direct say in how we scale as the platform grows. This is hands-on infrastructure work: capacity planning, distributed systems design, and deployment tooling engineers rely on every day. What you'll do - Architect and scale our AWS and Kubernetes infrastructure as usage and team size grow - Design and implement zero-downtime deployment strategies with automated rollback - Build and maintain the Terraform modules the rest of engineering relies on for infrastructure changes - Lead capacity planning and cost management across production systems - Build observability into our services with Honeycomb, Sentry, and CloudWatch so issues surface before they become incidents - Manage CI/CD pipelines in GitHub Actions that let engineers ship with confidence - Own networking and traffic routing decisions, including load balancing and DNS What we're looking for - 7+ years in DevOps, Site Reliability Engineering, or Infrastructure Engineering, with deep AWS expertise (EKS, VPC, S3, RDS/Aurora) and fluency in the Well-Architected framework - Has run Kubernetes in production, including cluster administration and multi-environment management, and writes Terraform at scale (module development, state management) - Has deployed CNCF tooling in production, like Helm and Karpenter, and built CI/CD pipelines in GitHub Actions - Has used observability tools (Honeycomb, Sentry, CloudWatch) to diagnose production issues, and implemented zero-downtime deployments with automated rollbacks - Strong networking fundamentals (TCP/IP, DNS, load balancing, traffic routing) and experience managing async job processing and message queues at scale - Scripts and automates in Python, Go, Bash, or similar, with a track record of capacity planning, performance tuning, and cost management Nice to have - Has designed alerting strategies and on-call runbooks for production services, and built dashboards for real-time service health - Background in database monitoring for RDS Postgres, plus experience building custom metrics and instrumentation for application-specific insight - Experience profiling applications to find performance bottlenecks, and working with Cloudflare WAF (origin shielding, firewall rules) - Familiarity with Okta SSO integration and experience designing RBAC across AWS and Kubernetes Why you'll love working here - Incredible ownership. We have the reach of a large company and the team size of a startup, so you'll own a large part of what we build. Customers like T-Mobile, LinkedIn, HubSpot, New York Life and...
Listing freshness
CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.