SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 175,000 - 240,000 / annual
Edison Scientific is building AI scientist agents to accelerate scientific discovery and medicine development. As a Senior Infrastructure Engineer (Member of Technical Staff), you will design, scale, and operate the core platform infrastructure powering autonomous scientific research.
This role focuses on technical ownership and leverage—understanding how complex systems interact, making sound architectural tradeoffs, and building foundations that enable teams and science to move faster. You'll own the infrastructure foundation that AI agents performing long-running scientific research depend on, which demands resilient scheduling, lifecycle management, and resource orchestration far beyond typical cloud-native workloads.
Key responsibilities include:
- Build, maintain, and improve Kubernetes-based infrastructure supporting agent workloads, research services, and internal platforms
- Contribute to cluster scaling, resource management, and operational improvements as platform usage grows
- Develop and maintain infrastructure tooling, automation, and deployment workflows to improve reliability and developer productivity
- Implement monitoring, observability, and alerting systems to ensure platform health and performance
- Support storage, networking, and security initiatives within Kubernetes environments
- Troubleshoot infrastructure and production issues across distributed systems; participate in incident response
- Collaborate with backend, ML, and research teams to understand workload requirements and implement reliable solutions
- Contribute to infrastructure best practices, documentation, and operational processes
You'll work with a team of scientists and engineers from leading institutions across biology, physics, chemistry, and AI in a fast-moving, mission-driven culture.
REQUIREMENTS:
- 6+ years of software, infrastructure, platform, or DevOps engineering experience
- Experience working with Kubernetes in development or production environments
- Familiarity with cloud platforms (AWS, GCP, or Azure)
- Proficiency in at least one programming language (Python, Go, Java, or TypeScript)
- Experience with infrastructure-as-code tools (Terraform, Pulumi, or similar)
- Understanding of containerized applications, networking fundamentals, and distributed systems concepts
- Experience using CI/CD pipelines, version control systems, and automated testing practices
- Ability to work independently while collaborating effectively across teams
- Curiosity, strong problem-solving skills, and desire to learn new technologies
PREFERRED:
- Security experience with gvisor/kata containers, snapshots