CronJobs

devops-sre jobs

SWE - Distributed

Achira · San Francisco Office

hybridunknown$164,638–$259,000Posted Oct 7, 2025KubernetesAWSGCPAzureRayPyTorchTensorFlowDistributed Systems

Apply on the employer site

About this role

**SWE - Distributed** **Why Achira** Join a world-class team of scientists, ML researchers, and engineers reshaping the future of drug discovery. Work on cutting-edge ML infrastructure at frontier scale with massive compute, data, and ambition. Own impactful work end-to-end from ideation to deployment. **About the Role** Architect and build infrastructure for ML data generation pipelines, model training, and fine-tuning workflows across large-scale distributed systems. You'll ensure compute clusters are efficient, observable, cost-effective, and reliable while pushing the boundaries of ML development. **What You'll Do** • Design and optimize distributed compute infrastructure for ML workloads • Improve cluster observability, scheduling, and resource utilization • Research and implement cost-efficient compute solutions • Develop monitoring and performance tuning tools • Collaborate with ML engineers to accelerate pipelines • Stay current with distributed computing technologies (Ray, Kubernetes, Spark, Slurm) **About You** • Experience building/working with distributed frameworks (Ray, Dask, Celery) • Strong grasp of parallel computing, job scheduling, and resource management • Proficient at identifying and resolving performance issues in distributed systems • Experience with cloud platforms (AWS, GCP, Azure) and orchestration (Kubernetes, Slurm) • Familiar with ML frameworks (PyTorch, TensorFlow, JAX) and MLOps best practices **Eligibility** All persons hired must verify identity and eligibility to work in the United States.

Listing freshness

CronJobs last confirmed this listing 6d ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord