SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Base Power is building America's next-generation power infrastructure through distributed battery networks that transform the centralized grid into a resilient, abundant system. The company is now launching a new business line: a distributed GPU fleet that attaches datacenter-grade compute to its power network and serves it to the AI industry, from hosted hardware and bare-metal nodes to an OpenAI-compatible inference API.
As Server Architect, you'll be a founding-team member building the complete infrastructure for this distributed compute platform. Your responsibilities span multiple critical domains:
Fleet Control Plane: Design and build systems to inventory, activate, and recover thousands of remote GPU nodes deployed across distributed locations. This includes identity management, telemetry collection, and operator tooling.
Out-of-Band Provisioning: Own the provisioning and recovery infrastructure using Redfish/iDRAC automation, network boot, immutable OS images, and rescue paths that eliminate ad hoc SSH sessions.
Node Agent Development: Create a secure, durable local substrate that establishes trust with the cloud, reports state, receives work, and survives reboots and poor network conditions.
Networking & Connectivity: Design overlay networking and secure connectivity across consumer internet links, making complex distributed systems reliable and boring.
Compute-Power Integration: Connect compute dispatch with energy planning systems so nodes run when power is available and cheap, and throttle or drain when the home or grid needs capacity.
Hardware Collaboration: Work with hardware and deployment teams on enclosures, thermals, and the realities of operating servers outdoors in Texas summers.
You'll need 8+ years building systems software close to hardware—platform management, firmware, provisioning, fleet orchestration, or backend services operating physical machines. Deep server platform expertise (BMC/iDRAC, Redfish, IPMI, PXE, hardware inventory, remote recovery) is essential. Strong programming in C, C++, Rust, or Go with comfort moving between boot logs and distributed services is required. Experience operating fleets you couldn't walk up to (servers, network gear, vehicles, telecom, energy hardware) is critical. You should have good judgment about what to build versus adopt, and strong ownership mentality for a small team with real revenue targets.
Nice-to-haves include datacenter or cloud platform background (hypervisors, bare-metal clouds, provisioning at scale), GPU serving experience (vLLM, inference routing, KV-cache scheduling), compiler/toolchain/OS-image build depth, or exposure to power systems and energy markets. Energy experience is not required—the company will teach you the grid.