CronJobs

devops-sre jobs

Sr. Software Engineer, DevOps

Akasa · San Francisco

hybridsenior$180,000–$220,000Posted May 1, 2026PythonKubernetesTerraformAWSGCPAzureGrafanaPrometheus

Apply on the employer site

About this role

**About AKASA** AKASA builds the future of healthcare with AI—providing generative AI solutions for the healthcare revenue cycle. We help health systems streamline operations and capture the full patient clinical journey. We’ve raised $205M+ from investors including Andreessen Horowitz, BOND, and Costanoa Ventures. Our AI-native product suite has grown 20x+ since launching in 2024, and deployments have been recognized nationally for real-world GenAI impact in healthcare finance. **Role: Sr. Infrastructure Engineer (DevOps / Platform)** As a Sr. Infrastructure Engineer, you’ll partner with our Infrastructure and Platform teams to manage, improve, and scale the systems powering our products. Your focus is on making infrastructure reliable, observable, and easy to operate—prioritizing automation, operational excellence, and cross-functional collaboration. You’ll help build and maintain foundational infrastructure for our SaaS applications, including **Kubernetes**, **Terraform-managed cloud resources**, and **GitHub-based CI/CD pipelines**. While incident response is part of the role, the primary emphasis is proactive improvements: reducing operational toil, improving visibility, and enabling product teams to move fast with confidence. **Location / Work Style** - Office: **South San Francisco** - Remote support: available across teams - Expectation: local R&D teams come into the office **every Wednesday** for co-working days (this role is expected to attend). --- **What You’ll Do** - **Infrastructure Management:** Build, manage, and optimize infrastructure using **Terraform**, **GitHub CI/CD**, and **Kubernetes** - **Monitoring & Observability:** Create dashboards and alerts using **Grafana**, **Prometheus/Mimir**, **OpenSearch**, and **Sentry** (or similar) - **Automation & Reliability:** Replace manual/error-prone processes with automated, repeatable systems - **Production Troubleshooting:** Diagnose and resolve production issues across application and infrastructure layers - **Documentation:** Maintain runbooks, setup guides, and architecture diagrams to support operational maturity - **Collaboration:** Drive adoption of DevOps and infrastructure best practices across teams - **Scalability Planning:** Scale infrastructure and monitoring systems as demand grows - **Incident Participation:** Join an on-call rotation and support incident response as needed --- **Skills & Qualifications** - **Observability:** Experience with metrics/logs/traces using **Grafana**, **Prometheus/Mimir**, **OpenSearch**, **Sentry**, or similar - **Infrastructure as Code:** Proficiency with **Terraform**, **Kubernetes**, and containerization - **Programming:** **5+ years** of experience with **Python** - **Linux Systems:** Comfortable with Linux environments and writing shell scripts - **Communication:** Strong collaboration skills with a focus on asynchronous, written communication - **Documentation:** Commitment to clear, comprehensive documentation and process standardization - **Initiative:** Self-starter with a proactive approach to operational challenges - **Version Control:** Skilled with **Git/GitHub** workflows **Nice to Haves** - **Cloud:** AWS (preferred), GCP, or Azure - **Networking:** TCP/IP, DNS, routing, load balancing fundamentals - **Security:** Understanding of cloud/infrastructure security best practices - **Performance Tuning:** Experience tuning production application/infrastructure performance --- **What We Offer** - Flexible p

Listing freshness

CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord