SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a text-to-speech platform used by over 50 million people globally to convert PDFs, books, Google Docs, articles, and websites into audio. The company was named Chrome Extension of the Year by Google and received Apple's 2025 Design Award for Inclusivity. With ~200 employees distributed globally and no physical office, Speechify combines consumer accessibility with enterprise solutions.
You'll join the Data side of the AI team, responsible for all aspects of data collection supporting model training operations. The team builds high-quality datasets at petabyte scale with tight integration of infrastructure, engineering, and research.
Key responsibilities:
- Source and integrate new audio data into the ingestion pipeline
- Operate and extend cloud infrastructure (GCP, Terraform) for data ingestion
- Partner with research scientists to optimize the cost/throughput/quality frontier, delivering richer datasets at scale and lower cost
- Collaborate with the AI team and leadership to shape the dataset roadmap for next-generation consumer and enterprise products
Required qualifications:
- BS/MS/PhD in Computer Science or related field
- 5+ years of industry software development experience
- Proficiency with bash/Python scripting in Linux environments
- Professional experience with Docker and Infrastructure-as-Code (GCP preferred)
- Strong communication skills
Desirable:
- Web crawler or large-scale data processing workflow experience
- Ability to juggle multiple priorities and adapt to change
The role offers competitive salary ($140k–$200k base), equity, and bonus. Work on a product impacting millions with learning differences including dyslexia, ADD, low vision, and autism. Join a fast-growing, entrepreneurial team with hands-off management and asynchronous culture.