SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a 50M+ user text-to-speech platform that transforms reading into audio across iOS, Android, Mac, Chrome, and web. The company was named Chrome Extension of the Year by Google and received Apple's 2025 Design Award for Inclusivity. With ~200 employees globally in a fully distributed setting, Speechify is building next-generation AI/ML products at the intersection of audio and accessibility.
You'll join the Data side of the AI team, responsible for all aspects of data collection to support model training operations. The team builds high-quality datasets at petabyte-scale and low cost through tight integration of infrastructure, engineering, and research.
Key responsibilities:
- Source new audio data and integrate it into the ingestion pipeline
- Operate and extend cloud infrastructure (GCP, Terraform) for data ingestion
- Collaborate with research scientists to optimize the cost/throughput/quality frontier, delivering richer datasets at scale and lower cost
- Work with the AI team and leadership to shape the dataset roadmap powering next-generation consumer and enterprise products
You bring 5+ years of software development experience, proficiency with Python/bash in Linux, Docker, Infrastructure-as-Code, and hands-on experience with a major cloud provider (ideally GCP). Experience with web crawlers and large-scale data processing is a plus. You're scrappy, adaptable, and communicate clearly.
The role offers competitive compensation, equity, and the opportunity to impact millions of users with learning differences including dyslexia, ADD, low vision, and autism. You'll work in a fast-growing, entrepreneurial environment with a hands-off management approach and strong asynchronous culture.