SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 141,000 - 208,000 / annual
ClickHouse, a Forbes Cloud 100 company and leader in real-time analytics and observability, is hiring a Cloud Software Engineer for its Observability Platform team. The role sits at the intersection of distributed systems, cloud infrastructure, and production operations.
The Observability Platform and Internal Observability teams build and operate systems that process trillions of events per day at hundreds of millions of events per second. You will design, build, and operate distributed systems for telemetry ingestion, durable buffering, processing, storage, autoscaling, and service provisioning. Your work directly feeds into the platform and product delivered to customers including Meta, Tesla, Sony, and Cursor.
Key responsibilities include:
- Design and operate distributed systems handling massive telemetry scale
- Own reliability, performance, capacity, and cost-efficiency of telemetry pipelines and storage
- Participate in on-call rotation and drive root-cause fixes for production incidents
- Build software and automation to eliminate repetitive operational work
- Identify architectural bottlenecks and shape the roadmap for scaling
- Collaborate with product, infrastructure, and service teams across ClickHouse
- Contribute to architecture reviews and raise engineering quality
You should have 5+ years building and operating production systems at scale, strong Go proficiency, hands-on Kubernetes experience, infrastructure-as-code expertise (Terraform, Helm, Argo CD), and production cloud experience (AWS, GCP, or Azure). Telemetry systems experience (OpenTelemetry, Prometheus, Grafana) is required. Bonus qualifications include ClickHouse experience, high-throughput systems background, multi-tenant cloud services, infrastructure cost optimization, and TypeScript skills.
The role emphasizes ownership, pragmatic tradeoffs, clear communication in async environments, incremental delivery validated in production, and addressing root causes rather than symptoms. You'll work in a remote, distributed environment with a company that has grown ARR 250% YoY and recently closed a $400M Series D.