SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Thought Machine is building modern banking infrastructure to replace legacy technology in banks worldwide. The company has raised over £500m from top-tier investors including JPMorgan Chase, Standard Chartered, and Temasek, and operates across London, New York, Singapore, Sydney, and Lisbon with 550+ employees.
As a Senior Infrastructure Engineer, you will play a crucial role in designing, deploying, and maintaining the cloud-native platform that powers Thought Machine's core banking and payments technology. You'll work on a platform infrastructure team whose mission is to abstract multi-cloud complexity and reduce cognitive load for other developers, treating infrastructure as a first-class product.
Key responsibilities include:
- Scaling and optimizing the cloud-agnostic platform infrastructure (Vault) that powers the system globally
- Building intuitive control planes and automation to simplify orchestration and reduce complexity for internal engineers and bank SREs
- Designing robust, fault-tolerant architectures with comprehensive disaster recovery and business continuity planning
- Enhancing observability suites shared across the product ecosystem, providing clients with deep monitoring and insights
- Scaling and optimizing core databases and event-streaming platforms for maximum performance and cost efficiency
You'll work with cutting-edge infrastructure tooling and have the opportunity to influence how thousands of developers and bank operators interact with cloud infrastructure at scale.
REQUIREMENTS:
- Degree in Computer Science, Engineering, or similar technical field
- 5+ years of hands-on software engineering experience building and scaling infrastructure platforms (not just maintaining them)
- Strong proficiency in production-ready platform tooling using Python or Golang
- Deep understanding of container and orchestration internals (Kubernetes, Docker)
- Hands-on experience extending and integrating open-source tools (Prometheus, OpenTelemetry, Grafana)
- Proven ability to package and scale developer-facing infrastructure utilities with focus on system reliability
DESIRABLE:
- Experience building cloud-agnostic solutions across AWS, Azure, and GCP
- Technical depth in databases or event-streaming platforms (Kafka, Postgres, DuckDB)
- Track record of troubleshooting, extending, or optimizing infrastructure tools professionally