SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Komodo Health is seeking a Senior Infrastructure Engineer to join its Infrastructure team, which builds, operates, and continuously improves the cloud-native stack and foundational systems powering the company's data ecosystem. The role focuses on keeping Komodo's AWS cloud infrastructure and Snowflake data platform reliable, secure, and cost-efficient.
Key responsibilities include operating and optimizing Snowflake in production—monitoring warehouse consumption and cost, administering RBAC and metadata tagging, and triaging query-performance issues. You will build and operate AWS and Kubernetes infrastructure as code using Terraform, deployed via GitOps with ArgoCD. You'll maintain and improve CI/CD pipelines using GitHub Actions, drive infrastructure optimization across performance, scalability, and cost through FinOps practices, and harden cloud infrastructure through IAM, RBAC, networking, and secrets management aligned with HIPAA and SOC2 compliance requirements.
Additional responsibilities include building and maintaining centralized observability (metrics, logs, alerting) and participating in a shared US-business-hours on-call rotation while collaborating closely with developer, security, and data teams.
Required qualifications include hands-on Snowflake or comparable cloud data warehouse experience in production (SQL, warehouse operations, query-performance fundamentals), hands-on AWS production experience with attention to security and cost, Infrastructure-as-Code proficiency with Terraform, containerization and Kubernetes experience (Docker, k8s), scripting and automation skills in Python, Go, or Bash, and CI/CD pipeline experience with tools like GitHub Actions, Jenkins, or CircleCI. Experience with AI-assisted engineering tools is also expected.
Within the first 12 months, you will have operated and optimized the Snowflake environment, contributed to measurable cloud and Snowflake cost reduction, delivered and operated production AWS and Kubernetes infrastructure, strengthened security and compliance posture, and strengthened observability and on-call readiness.