SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Tabby is a fintech company reshaping how people shop, earn, and save. With over 25 million users and $18 billion in annual transaction volume, Tabby is the largest and fastest-growing fintech in the GCC region, valued at $6.5 billion.
You will join the Infrastructure team as a Senior Database Administrator, owning the reliability and performance of Tabby's production ClickHouse clusters. This is a high-impact role in a high-growth environment working alongside a world-class remote engineering team across 20+ countries.
Key responsibilities:
- Own reliability and performance of production ClickHouse clusters (self-hosted on Kubernetes and managed services)
- Design standard architecture and tooling to scale to dozens of multi-node clusters without proportional operational overhead
- Manage cluster lifecycle as code: provisioning, configuration, scaling, upgrades, and migrations
- Build monitoring and alerting to catch problems before users notice them
- Own backups, disaster recovery, and regularly test restores
- Ensure cluster security by default
- Plan capacity and control infrastructure costs
- Serve as the go-to ClickHouse expert for developers, DevOps engineers, other DBAs, and leadership on schema design, data ingestion, and query performance
- Handle routine database requests and automate them
- Respond to production emergencies outside business hours, lead root-cause analysis, and contribute to postmortems
- Maintain documentation, runbooks, and share knowledge with the team
Requirements:
- Deep, DBA-level production knowledge of ClickHouse: understand how it stores, merges, replicates, and queries data; strong SQL skills including complex analytical query optimization and schema design review
- Ability to design architecture and automation for dozens of multi-node clusters; experience running large fleets of databases (dozens or hundreds of instances) is a strong advantage
- Production experience running and maintaining stateful workloads on Kubernetes well beyond initial deployment
- Production use of a Kubernetes operator for databases (CloudNativePG, Zalando Postgres Operator, Altinity, ClickHouse operator, or equivalent)
- Experience with version upgrades, backup/restore, and migrating large datasets between clusters or regions with zero or minimal downtime
- Strong troubleshooting skills across the full stack: queries, schema design, Kubernetes, Linux, storage, and networking
- Proactive approach to database health: built monitoring and alerting with modern observability stacks and act on early warning signs
- Practical database security experience: encryption in transit and at rest, access control, secrets management
- Infrastructure as code with mainstream tools plus at least one programming or scripting language for automation
- Hands-on experience with cloud infrastructure on major providers (not just platform-level usage)
- Clear written and spoken communication with engineers and leadership on architecture, performance, incidents, risks, and costs; ability to take problems from investigation to production fix independently
Nice to have:
- Hands-on work on Kubernetes database operator source code as developer or contributor
- Practical use of AI tools for database observability or automating routine tasks
- Background in highly regulated environments
- Kubernetes or cloud certifications