SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a 50M+ user text-to-speech platform that transforms reading into audio across iOS, Android, Mac, Chrome, and web. The company was named Chrome Extension of the Year by Google and won Apple's 2025 Design Award for Inclusivity. With ~200 employees globally in a fully distributed setting, Speechify operates at the intersection of AI and audio accessibility.
You'll join the Data side of the AI team, responsible for all aspects of data collection supporting model training operations. The team builds high-quality datasets at petabyte scale and low cost through tight integration of infrastructure, engineering, and research.
Key responsibilities include:
- Sourcing new audio data and integrating it into the ingestion pipeline
- Operating and extending cloud infrastructure (GCP, Terraform) for data ingestion
- Collaborating with research scientists to optimize the cost/throughput/quality frontier, delivering richer datasets at scale and lower cost
- Working with the AI team and leadership to shape the dataset roadmap powering next-generation consumer and enterprise products
You should have a BS/MS/PhD in Computer Science or related field, 5+ years of software development experience, proficiency with bash/Python in Linux, Docker, Infrastructure-as-Code, and experience with at least one major cloud provider (GCP preferred). Web crawlers and large-scale data processing experience is a plus. The role values adaptability, strong communication, and scrappiness in a fast-moving environment.
Speechify's mission is ensuring reading is never a barrier to learning, with products directly impacting people with dyslexia, ADD, low vision, concussions, autism, and other learning differences.