Software Engineer, ML Platform
Cursor · San Francisco
About this role
**Software Engineer, ML Platform** **About the Role** Build infrastructure that turns product usage into better models and keeps research moving fast on large GPU fleets. ML Platform has four teams: • **Telemetry** - Collection and serving path for product data • **ML Data Platform** - Shared environments and pipeline substrate • **Observability** - Tools for researchers to start, watch, and debug runs • **ML DevX and Systems** - Path from idea to trusted runs on research fleet **What You'll Do** • Design, build, and operate core platform systems for ML researchers • Partner with research to turn pain points into durable infrastructure • Own reliability, performance, and developer experience • Ship iteratively with high ownership in a flat environment **You May Be a Fit If** • Strong background in systems/infrastructure software engineering • Experience owning production distributed systems at scale • Comfortable with Linux, cloud/bare metal, and modern orchestration (Kubernetes, Ray, etc.) • Enjoy working closely with ML researchers and product engineers • Thrive with high ownership and short feedback loops **Especially Strong Backgrounds** • Telemetry: event ingestion, analytics pipelines, OpenTelemetry, data APIs • Data Platform: data frameworks, Spark/Flink/Ray, ML infrastructure • Observability: experiment monitoring, debug tooling, observability UX • ML DevX: GPU scheduling, job queues, cluster health, compute experience **Location:** In-person in San Francisco, Palo Alto, or Manhattan
Listing freshness
CronJobs last confirmed this listing 19h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.