SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 30,000 - 110,000 / annual
Speechify is a text-to-speech platform used by over 50 million people globally to convert PDFs, books, Google Docs, articles, and websites into audio. The company was recently named Chrome Extension of the Year by Google and received Apple's 2025 Design Award for Inclusivity. With ~200 employees distributed globally and no physical offices, Speechify operates a fully remote, asynchronous culture.
You'll join the Data side of the AI team, responsible for all aspects of data collection supporting model training operations. The team builds high-quality datasets at petabyte scale through tight integration of infrastructure, engineering, and research. This is a hands-on role where you'll directly impact the next generation of Speechify's consumer and enterprise AI products.
Key responsibilities include: sourcing new audio data and integrating it into ingestion pipelines; operating and extending cloud infrastructure (GCP, Terraform-managed) for data pipelines; collaborating with research scientists to optimize the cost/throughput/quality frontier; and working with leadership to shape the AI team's dataset roadmap.
You'll need a BS/MS/PhD in Computer Science or related field, 5+ years of software development experience, proficiency in bash/Python scripting on Linux, hands-on experience with Docker and Infrastructure-as-Code, and professional experience with a major cloud provider (GCP preferred). Experience with web crawlers and large-scale data processing workflows is a plus. The ideal candidate is scrappy, adaptable, and communicates clearly in writing and verbally.
Speechify offers competitive salaries, stock options, a laid-back atmosphere, and the opportunity to work on a product that directly impacts millions of people with learning differences including dyslexia, ADD, low vision, and autism.