SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Kikoff is a profitable, pre-IPO fintech company on a mission to empower financial security at scale. The infrastructure team builds systems that enable engineering teams to move quickly without sacrificing reliability, security, or cost discipline.
As a Staff Infrastructure Engineer, you will own high-leverage infrastructure problems from design through production operation. You will build internal products and paved paths that turn ambiguous problems into durable systems. Your work should improve delivery speed and reliability, reduce cost, and reduce operational toil for product teams.
This is a hands-on Staff IC role. You will foster relationships, write code, review designs, make architecture decisions, lead through incidents, and help engineers solve problems outside the runbook. You will have a primary area of depth and enough range to follow production problems across infrastructure boundaries.
Key responsibilities include:
**Build and Own Critical Platform Systems**: Design and implement self-service infrastructure on AWS using reusable code and infrastructure-as-code patterns (Pulumi, HCL, TypeScript). Own the systems you build in production, including reliability, security, capacity, cost, upgrades, incidents, and recovery. Give critical services an SLO, actionable alerts, dashboards, runbooks, and tested recovery paths. Automate recurring operational work and eliminate failure-prone manual steps.
**Run Infrastructure as a Product**: Work directly with engineers to turn recurring friction into paved paths, self-service tools, and automated workflows. Measure outcomes for internal customers through adoption, developer feedback, delivery speed, reliability, cost, toil, and on-call load. Use the fastest responsible path when a team is blocked, then turn friction into automation or durable platform capability.
**Set Technical Direction**: Set technical direction for ambiguous infrastructure work from problem framing through implementation, rollout, and production ownership. Make clear trade-offs among delivery speed, reliability, security, cost, and long-term operational complexity. Partner across Infrastructure, Security, Data, and Product Engineering to build security and compliance controls into normal workflows.
**Raise the Engineering Bar**: Review code and designs, challenge weak assumptions, help other engineers make better technical decisions, and uphold high standards for reliability, testing, safe deployments, security, and maintainability. Lead through incidents and unfamiliar failure modes. Create standards, reusable patterns, and durable documentation.
You have 7+ years of infrastructure, platform, or software engineering experience, or an equivalent record of Staff-level technical impact. You have sustained ownership of a consequential production system, handled incidents and unfamiliar failure modes, and improved the system afterward. Strong coding skills in TypeScript, Python, Go, Ruby, or similar languages. Production experience with AWS, infrastructure as code, containers, CI/CD, observability, and security boundaries. Experienced generalist judgment with depth in at least one infrastructure domain.