AI Inference Core - Senior SW Engineer for Platform & DevOps
Cerebras · Sunnyvale, CA
About this role
**Cerebras Systems — AI Inference Core (Senior SW Engineer for Platform & DevOps)** Cerebras builds the world’s largest AI chip—56x larger than GPUs—enabling industry-leading training and inference speeds (10x faster than GPU-based hyperscale cloud inference). This performance unlocks real-time iteration and more agentic computation for AI applications. **About the Team** The Core Infrastructure team builds the software systems that power engineering workflows across Cerebras. You’ll help coordinate complex work across machines, clusters, development environments, and hardware systems—building orchestration frameworks, execution engines, scheduling systems, test infrastructure, developer tools, and reusable software platforms. **About the Role** We’re hiring a Software Engineer to build and operate the platform layer behind Cerebras engineering infrastructure. This is an engineering-focused infrastructure role (not primarily ticket-driven operations). You’ll work on: - CI/CD systems - Kubernetes and deployment automation - Cloud + on-prem infrastructure - Developer environments - Artifact management - Observability You’ll automate repeated work, debug failures across system boundaries, and turn operational problems into durable platform improvements. **Responsibilities** - Design, build, and maintain CI/CD systems for build, test, integration, qualification, and release workflows - Build and operate Kubernetes-based platforms and services for engineering teams - Develop deployment systems, internal tools, and self-service workflows that make changes repeatable, reviewable, and safe - Improve reliability, capacity, performance, cost efficiency, monitoring, and operational readiness - Debug issues across CI pipelines, Kubernetes workloads, networking, storage, authentication, operating systems, and distributed applications - Perform root-cause analysis and implement lasting fixes - Partner with software, IT, security, networking, release, and developer-productivity teams **Skills & Qualifications** - 5+ years of experience in platform engineering, DevOps, infrastructure engineering, SRE, or software engineering - Hands-on experience with CI/CD and automated software-delivery workflows - Experience deploying and operating services with Kubernetes/containerized environments - Experience with a major cloud platform (preferably AWS) and programmatic infrastructure provisioning - Strong Linux/Unix fundamentals - Networking concepts: DNS, routing, load balancing, proxies, ports, TLS, service connectivity - Proficiency in Python, Shell, or another language for infrastructure automation/tooling - Experience with monitoring, logging, alerting, dashboards, and incident investigation - Strong debugging/problem-solving across applications, infrastructure, networking, and OS **Preferred Qualifications** - Terraform (infrastructure-as-code) - Kubernetes controllers/operators, custom resources, Helm, Argo CD, or similar platform technologies - Artifact repositories/package registries/build caches/software distribution infrastructure - Build systems, dependency management, reproducible builds - Hybrid environments (cloud + on-prem + specialized hardware) - IAM, secrets, certificates, TLS/mTLS - Internal developer platforms / self-service infrastructure products - BS/MS in CS or related field (or equivalent practical experience) **Why Join Cerebras** People who are serious about software make their own hardware. At Cerebras, you’ll work on a brea
Listing freshness
CronJobs last confirmed this listing 2h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.