Senior Systems Engineer, Workers AI
Cloudflare · In-Office
About this role
## About Cloudflare At Cloudflare, we’re on a mission to help build a better Internet. We protect and accelerate Internet applications online—without adding hardware, installing software, or changing a line of code. We’re looking for builders who spot “normalized” problems and use AI-native curiosity to create solutions with the latest tools. Our culture is built on iteration: ship faster today to make it better tomorrow, and share improvements across the team. ## Location **Austin, TX or London, UK (Hybrid)** ## About the role You’ll design and build the core infrastructure that powers **AI inference across Cloudflare’s global network**—including real-time voice, frontier open LLMs, and customer-deployed models running on a heterogeneous fleet of GPUs and next-generation accelerators across hundreds of cities. You’ll work with AI/ML engineers, hardware partners, and Cloudflare product teams to solve hard problems in **distributed systems** and **high-performance computing**, such as: - Sub-second model cold starts - Multi-accelerator workload scheduling - Efficient KV cache management - A model deployment platform serving both Cloudflare and customer “bring-your-own-model” workloads ## Responsibilities - **Platform Architecture:** Develop and maintain core components of a serverless inference platform to ensure high availability and scalability. - **Optimization & Performance:** Improve model scheduling efficiency and resource utilization; enhance request routing to reduce end-user latency. - **System Reliability:** Identify and mitigate systemic risks to measurably improve reliability and resilience. - **Observability:** Expand/refine metrics, logging, and tracing; fine-tune alerts to proactively detect and resolve production issues. - **Technical Leadership:** Lead complex cross-functional projects from concept and design through deployment and operationalization. - **Mentorship:** Mentor junior engineers and help cultivate a collaborative engineering culture. ## Desirable skills, knowledge & experience - Proven systems engineering experience, focused on **distributed, high-performance systems** - Expert proficiency in **Rust**, especially in asynchronous environments - Deep hands-on understanding of networking and application protocols (e.g., **TCP, HTTP, WebSocket**) - Experience scaling systems and optimizing performance (e.g., **load balancing** and **caching**) ## Bonus points - Experience with container orchestration platforms, specifically **Kubernetes** and/or **Nomad** - Familiarity with large-scale inference serving challenges (e.g., **LLMs** and diffusion models) ## What makes Cloudflare special? We’re not just ambitious—we’re ambitious with a mission to protect the free and open Internet. Examples include: - **Project Galileo** - **Athenian Project** - **1.1.1.1** (privacy-centric public DNS resolver; Cloudflare does not store client IP addresses) ## Notes - Applicants progressing to the offer stage may be asked to attend an in-person interview at a Cloudflare office/hub. - This role may involve access to information protected under U.S. export control laws. - Cloudflare is an equal opportunity employer and provides reasonable accommodations for qualified individuals with disabilities.
Listing freshness
CronJobs last confirmed this listing 15h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.