SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Mirantis is seeking a Senior Software Engineer to join its cloud-native infrastructure team building next-generation bare-metal orchestration platforms for AI workloads. The team operates on a T-shaped engineering model where each engineer brings deep expertise in a primary domain while maintaining practical fluency across neighboring systems. Development is fundamentally based on open-source projects, with primary control plane software written in Rust and Go, leveraging Kubernetes operators and an OpenTelemetry-first operational culture.
In this role, you will drive the development of control planes powering high-density, high-throughput infrastructure for AI training and inference workloads. Core responsibilities include:
• Design, build, and maintain production-grade control plane microservices, custom Kubernetes operators, and robust reconciliation engines in Rust and Go
• Author clear Functional and Non-Functional Requirements, architectural specs, system sequence diagrams, and clean gRPC/Protobuf and REST API schemas
• Develop custom integration modules for modern network operating systems (SONiC, NVUE, Cumulus) and DPU hardware platforms
• Instrument services end-to-end using OpenTelemetry traces and metrics, performing systematic root-cause analysis across polyglot distributed systems
• Participate in rigorous, review-gated pull request workflows across Rust, Go, and SQL codebases while ensuring high test coverage via mock-driven testing
You will work with an established Silicon Valley leader in cloud infrastructure, collaborating with exceptionally talented colleagues on cutting-edge, open-source innovation serving Fortune 500 and Global 2000 customers.
REQUIREMENTS:
Technical Qualifications:
• Primary mastery of Rust (Tokio async runtime, Tonic, Axum, sqlx) and strong proficiency in Go
• Advanced SQL/PostgreSQL fluency, gRPC/Protobuf contract design, and schema evolution
• Strong background in async concurrency models, lock-free patterns, distributed state handling, and mock-driven testing discipline
• Expertise in BGP, MP-BGP, EVPN, VXLAN, L3VNI, route targets, and route server design
• Hands-on experience with SONiC, Cumulus Linux, and NVUE
• Deep knowledge of NVIDIA DOCA, Host-Based Networking (HBN), BlueField DPU architectures, and the DOCA Platform Framework operator model
• Understanding of InfiniBand fabrics, NVLink/NMX-M partitioning, and RoCEv2
• Advanced grasp of Linux netlink, network namespaces, routing tables, and internal DHCP/DNS service implementations
• Familiarity with CNI plugins (Calico, Cilium); OVS and DPDK experience is a plus
• Expertise in PXE/iPXE, UEFI, Secure Boot, measured boot, and TPM attestation
• Proficiency in Redfish, IPMI, and BMC abstractions across heterogeneous hardware platforms (Dell, Lenovo, NVIDIA reference hardware)
• In-depth knowledge of boot chains, systemd, initramfs, BIOS configuration matrices
• Practical experience with PKI, X.509 certificates, TLS, SPIFFE/SVID, Vault, KMS, Keycloak (OAuth2/JWT), and RBAC
• Experience with KubeVirt/KVM
• Experience writing custom controllers/operators using kube-rs or controller-runtime with CRD-driven reconciliation loop patterns
Universal Requirements:
• Commitment to test-driven design, writing highly testable code with mock interfaces for external hardware and network dependencies
• Proven track record of writing crisp architectural specs and engaging in constructive, cross-functional design reviews
Educational & Experience Baseline:
• Bachelor's or Master's degree in Computer Science, Computer Engineering, or equivalent practical experience
• 5+ years of software engineering experience in cloud-native platforms, systems programming, networking, or infrastructure automation
• Demonstrated participation or maintainership in open-source systems projects is a plus