SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 172,000 - 209,000 / annual
Crusoe is a vertically integrated AI infrastructure company building energy-efficient cloud systems to power the world's most ambitious AI workloads. The company owns and operates the full stack—from electrons to tokens—with an energy-first approach to solve AI compute bottlenecks.
You will architect, design, and develop cloud infrastructure management systems and platforms that enable efficient planning, monitoring, deployment, and operation of Crusoe's AI-first cloud. This is a hands-on role where you'll deliver end-to-end use cases and workflows, directly contributing to key business revenue metrics.
Key responsibilities include:
- Collaborating across teams to design and implement physical infrastructure management software systems, availability platforms, and frameworks for customers hosted on Crusoe's AI infrastructure
- Contributing to reliability, scalability, and security of systems and platforms
- Developing workflows that drive efficiency and meet business objectives
- Designing and implementing high-performing, highly available cloud architectures optimized for performance and cost-effectiveness
- Streamlining cloud deployment, configuration management, and operations through effective platforms, interfaces, and automation tooling
- Contributing to platform evolution through cross-functional collaboration
- Mentoring junior engineers and contributing to team growth
You'll work in a fast-paced startup environment focused on building energy-efficient, scalable AI infrastructure with a commitment to sustainability and clean energy innovation.
REQUIREMENTS:
- Bachelor's degree in Computer Science or Software Engineering
- 5+ years of relevant industry experience
- 5+ years of experience building and operating distributed systems at scale
- Proven experience building reliable, scalable, efficient, and secure cloud platforms and systems, with production operations experience
- Fluency in one or more of: Go, Rust, Java, or C++
- Collaborative mindset working with development and operations teams
- Solid understanding of cloud security best practices and ability to implement secure configurations
- Excellent troubleshooting and problem-solving skills for complex infrastructure challenges
- Strong written and verbal communication skills
BONUS:
- Hands-on experience deploying, managing, and troubleshooting Kubernetes clusters
- Experience in fast-paced startup environments
- Passion for energy-efficient, scalable AI infrastructure
- Enthusiasm for sustainability and clean energy innovation