SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Tabby is a fintech platform that enables flexible payment solutions, serving over 25 million users and processing $18 billion in annual transaction volume. The company is valued at $6.5 billion and operates across the GCC region and globally.
You will join the Infrastructure team as a Senior Database Administrator responsible for the reliability, performance, and security of Tabby's production ClickHouse clusters. This is a high-impact role in a high-growth environment working with a distributed engineering team across 20+ countries.
Key responsibilities include:
- Own the reliability and performance of production ClickHouse clusters, both self-hosted on Kubernetes and managed services
- Design standardized architecture and tooling to scale to dozens of multi-node clusters without proportional operational overhead
- Manage the complete cluster lifecycle as code: provisioning, configuration, scaling, upgrades, and migrations
- Build comprehensive monitoring and alerting to detect issues proactively
- Own backup and disaster recovery strategies, with regular restore testing
- Ensure security by default across all clusters
- Plan capacity and optimize infrastructure costs
- Serve as the primary ClickHouse expert for developers, DevOps engineers, other DBAs, and leadership on schema design, data ingestion, and query performance
- Automate routine database requests and operational tasks
- Respond to production emergencies outside business hours and lead root-cause analysis and postmortems
- Maintain documentation, runbooks, and knowledge sharing with the team
Requirements:
- Deep, production-level DBA knowledge of ClickHouse, including understanding of data storage, merging, replication, and query mechanics
- Strong SQL skills, including optimization of complex analytical queries and schema design review
- Experience designing architecture and automation for large fleets of databases (dozens to hundreds of instances); engine type is flexible
- Hands-on production experience running and maintaining stateful workloads on Kubernetes beyond initial deployment
- Production use of a Kubernetes operator for database management (CloudNativePG, Zalando Postgres Operator, Altinity, official ClickHouse operator, or equivalent)
- Experience with version upgrades, backup/restore, and zero or minimal-downtime migrations of large datasets between clusters or regions
- Strong troubleshooting skills across the full stack: queries, schema, Kubernetes, Linux, storage, and networking
- Proactive approach to database health with experience building monitoring and alerting using modern observability tools
- Practical database security knowledge: encryption in transit and at rest, access control, secrets management
- Infrastructure as code expertise with at least one programming or scripting language for automation
- Hands-on cloud infrastructure experience with a major provider (AWS, GCP, Azure)
- Clear written and spoken communication with engineers and leadership
Nice to have:
- Hands-on work on Kubernetes database operator source code as a developer or contributor
- Practical use of AI tools for database observability or operational automation
- Background in highly regulated environments
- Kubernetes or cloud certifications