CronJobs

devops-sre jobs

Server Architect

Base Power · Austin, TX

onsiteseniorPosted Aug 29, 2026GoRustCC++RedfishiDRACIPMIPXE

Apply on the employer site

About this role

**ABOUT BASE** Base is America’s next-generation power company—rebuilding the foundation of modern civilization by deploying a vast network of distributed batteries that transforms today’s fragile, centralized grid into a resilient and abundant system. **ABOUT THE ROLE** The hardest problem in AI compute today isn’t chips. It’s power—permitted, interconnected, and available now. Base already has it at thousands of sites, with more coming online every day. This is a **founding-team** role to help build a **Base-built, Base-operated distributed GPU fleet** that attaches datacenter-grade compute to our power network and serves it to the AI industry—from hosted hardware and bare-metal nodes to an **OpenAI-compatible inference API** running on our own machines. You’ll help build the whole system: - The **control plane** (inventories, activates, recovers thousands of remote nodes) - The **provisioning path** (flashes servers over **Redfish/iDRAC** from a thousand miles away) - The **node agent** (keeps nodes trusted, observed, and working) - The **dispatch layer** (decides when compute runs, throttles, or drains—together with energy planning) --- **WHAT YOU’LL DO** - Build the **fleet control plane**: inventory, identity, activation, telemetry, and operator tooling for a growing network of GPU nodes in the field - Own **out-of-band provisioning and recovery** - Implement **Redfish/iDRAC automation**, network boot, immutable OS images, and rescue paths that eliminate ad hoc SSH sessions - Develop the **node agent**: a secure, durable local substrate that establishes trust with the cloud, reports state, receives work, and survives reboots and bad networks - Design **overlay networking** and secure connectivity across consumer internet links—make it boring - Connect **compute to power**: integrate dispatch with energy planning so nodes run when power is available/cheap and back off when needed - Partner with hardware/deployment teams on enclosures, thermals, and the realities of running servers outdoors (e.g., Texas summers) --- **WHAT YOU’LL BRING** - **8+ years** building systems software close to hardware: platform management, firmware, provisioning, fleet orchestration, or backend services operating physical machines - Deep experience with **server platforms** and remote management: - **BMC/iDRAC**, **Redfish** or **IPMI** - **PXE/virtual-media boot**, hardware inventory, remote recovery - Strong **C, C++, Rust, or Go** skills; comfort moving between a boot log and a distributed service in the same day - Experience operating fleets you couldn’t walk up to: servers, network gear, vehicles, telecom, or energy hardware - Judgment about what to build vs. adopt (this project leans on standards and supply chains, not one-off inventions) - **Ownership**: small team, new business line, real revenue targets—set patterns others follow --- **NICE TO HAVES** - Datacenter/cloud platform background (hypervisors, bare-metal clouds, provisioning at scale) - GPU serving experience (vLLM, inference routing, KV-cache-aware scheduling, or GPU fleet operations) - Compiler/toolchain/OS-image build depth - Exposure to power systems or energy markets (you don’t need energy experience—we’ll teach you) --- **ABOUT THE TEAM** Most infrastructure roles scale someone else’s platform. This one starts a new kind of datacenter—distributed across thousands of homes, powered by batteries we build, running compute we operate end to end. The problems are unsolv

Listing freshness

CronJobs last confirmed this listing 2h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord