SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Storage Systems Engineer - remote in the US

Mirantis - Remote - Remote - posted 2026-09-11

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Mirantis is seeking a Senior Storage Systems Engineer to design, deploy, integrate, and operate high-performance storage systems for GPU-accelerated compute and AI platforms. You will own the storage layer where Kubernetes meets bare metal, standing up NFS-based high-performance storage, wiring it into clusters via Container Storage Interface (CSI), and tuning it to keep data flowing to GPU workloads at scale. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack. In this role, you will treat storage as infrastructure to be automated, observed, and tuned—not hand-managed. You are fluent in Kubernetes storage, deeply versed in Linux storage and networking fundamentals down to the kernel and NFS-client layer, and know how to make high-performance NAS actually perform under demanding workloads. You reach for infrastructure-as-code and GitOps by default, are self-directed in diagnosing performance and reliability issues end to end, set operational standards for others to follow, and communicate clearly across teams. Key responsibilities include: **Storage Integration & Operation:** Integrate NFS-based high-performance storage systems (e.g., VAST, Dell PowerScale) into Kubernetes clusters via CSI, storage classes, and persistent volumes. Tune the NFS data path—mount options, nconnect/RDMA, Linux client, and network settings—for high-throughput, low-latency GPU/AI workloads. Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle. **Linux Platform & System Integration:** Configure and optimize Linux systems for storage workloads, including driver setup, file system layout, network tuning, and kernel parameter optimization. Deliver storage integration for k0s-based Kubernetes via Cluster API (CAPI) and K0rdent management/child cluster topologies. Operate storage in fully disconnected (air-gapped) environments, including local artifact/mirror connectivity (Harbor) and PKI/TLS considerations. **Automation & Observability:** Automate storage provisioning and configuration with infrastructure-as-code (Terraform/OpenTofu) and GitOps pipelines (ArgoCD or Flux). Build monitoring, alerting, and observability for storage performance, capacity, and health. Diagnose and resolve performance, reliability, and scaling issues across the storage stack. **Requirements:** - 7+ years of experience in SRE or hardware/storage infrastructure operations - 5+ years of building/operating distributed production storage systems at scale - 7+ years of experience in Linux and Kubernetes storage fundamentals (NFS, CSI) - 1+ years of experience integrating with or building high-performance storage solutions (VAST, Weka, DDN, PowerScale) Baremetal hardware experience is a strong plus, but deep Linux storage knowledge is essential.

Similar roles