SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a text-to-speech platform used by over 50 million people globally to convert PDFs, books, Google Docs, articles, and websites into audio. The company was recently named Chrome Extension of the Year by Google and received Apple's 2025 Design Award for Inclusivity. With ~200 employees distributed globally and no physical office, Speechify combines engineering, AI research, and product development to serve users with learning differences and accessibility needs.
You'll join the Data side of Speechify's AI team, responsible for all aspects of data collection supporting model training operations. The team builds high-quality datasets at petabyte scale through tight integration of infrastructure, engineering, and research. This is a hands-on role where you'll be scrappy in sourcing new audio data and bringing it into the ingestion pipeline.
Key responsibilities include: operating and extending cloud infrastructure for the ingestion pipeline (currently GCP with Terraform), collaborating with research scientists to optimize the cost/throughput/quality frontier, and working with leadership to shape the AI team's dataset roadmap for next-generation consumer and enterprise products.
You'll need a BS/MS/PhD in Computer Science or related field, 5+ years of software development experience, proficiency with bash/Python in Linux, Docker and Infrastructure-as-Code expertise, and hands-on experience with at least one major cloud provider (GCP preferred). Experience with web crawlers and large-scale data processing workflows is a plus. The role demands strong communication, adaptability, and the ability to balance multiple priorities in a fast-moving environment.
Speechify offers competitive compensation, equity, a hands-off management approach, and the opportunity to impact millions of users while working at the intersection of AI and audio.