SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a 50M+ user text-to-speech platform that turns PDFs, books, articles, and websites into audio. The company has won Chrome Extension of the Year (Google) and 2025 Design Award for Inclusivity (Apple). With ~200 employees globally in a fully distributed setting, Speechify is building next-generation AI models for accessibility.
You'll join the Data side of the AI team, responsible for all aspects of data collection to support model training operations. The team builds high-quality datasets at petabyte-scale and low cost through tight integration of infrastructure, engineering, and research.
Key responsibilities:
- Source new audio data and integrate it into the ingestion pipeline
- Operate and extend cloud infrastructure for data ingestion (GCP, Terraform)
- Collaborate with research scientists to optimize the cost/throughput/quality frontier
- Work with AI leadership to shape the dataset roadmap for next-generation consumer and enterprise products
Required qualifications:
- BS/MS/PhD in Computer Science or related field
- 5+ years of software development experience
- Proficiency with bash/Python scripting in Linux
- Professional experience with Docker and Infrastructure-as-Code
- Experience with at least one major cloud provider (GCP preferred)
- Strong communication skills
Desirable: web crawlers, large-scale data processing workflows.
The role offers competitive salary ($140K–$200K base + bonus + equity), entrepreneurial environment, hands-off management, and the opportunity to impact millions of users with accessibility-focused AI.