SlipstreamJobsFresh Startup & VC-Backed Jobs

Staff Engineer

Graphcore - Austin, TX, United States - Hybrid - posted 2026-08-11

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Graphcore, a SoftBank Group company and leader in AI compute infrastructure, is seeking an experienced Staff Engineer to join the System Management team. This role sits at the critical intersection of hardware and customer workloads, responsible for building the foundational interfaces and infrastructure that enable internal and external customers to manage system state at scale. You will own software engineering efforts across the full software development lifecycle—from implementation and automated testing through integration and production readiness—for Graphcore's rack management solution. This includes configuring and testing new AI hardware and systems using continuous deployment and infrastructure-as-code practices in both internal and external datacenters. You'll drive ownership of critical infrastructure, collaborating across teams to resolve complex issues and maintain peak system performance alongside datacenter operations engineers. The System Management team, part of the Software Platform group, builds interfaces between hardware and AI software frameworks, providing solutions for public and private cloud deployments. As one of the first teams to work with pre-release hardware and software, you'll need comfort with unproven components and strong problem-solving capabilities. Essential qualifications include a bachelor's degree or equivalent, deep experience with RESTful API development, Kubernetes and container orchestration (Docker/Podman), production Kubernetes cluster management, Go programming, infrastructure-as-code tools (Terraform/OpenTofu, Ansible), Redfish for datacenter hardware management, Linux systems engineering, and Bash/Python scripting. You should be experienced with AGILE/SCRUM frameworks and capable of specifying, scoping, and detailing work plans. Desirable skills include experience with AI coding assistants, Kubernetes operator development, HPC environments (SLURM), and virtualized deployments. The ideal candidate thrives in fast-paced, loosely scoped environments, leads with action and decisiveness, and is versatile and hands-on.

Similar roles