SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a text-to-speech platform used by over 50 million people globally to convert PDFs, books, Google Docs, articles, and websites into audio. The company has been recognized by Google (Chrome Extension of the Year) and Apple (2025 Design Award for Inclusivity) and operates as a fully distributed, 100% remote organization with ~200 employees across engineering, AI research, and product teams.
You'll join the Data side of Speechify's AI team, responsible for all aspects of data collection and pipeline infrastructure supporting model training operations. The team builds high-quality datasets at petabyte scale while optimizing for cost through tight integration of infrastructure, engineering, and research.
Key responsibilities include: sourcing and integrating new audio data sources into the ingestion pipeline; operating and extending cloud infrastructure (GCP, Terraform-managed); collaborating with research scientists to optimize the cost/throughput/quality frontier; and working with leadership to define the AI team's dataset roadmap for next-generation consumer and enterprise products.
You should have a BS/MS/PhD in Computer Science or related field, 5+ years of software development experience, proficiency with bash/Python in Linux environments, hands-on experience with Docker and Infrastructure-as-Code, and professional experience with a major cloud provider (GCP preferred). Experience with web crawlers and large-scale data processing workflows is a plus. The role requires strong communication, adaptability, and comfort working in a fast-moving, asynchronous environment.
Compensation: $140,000–$200,000 base salary plus bonus and equity. The company emphasizes a hands-off management approach, entrepreneurial culture, and the opportunity to impact millions of users, particularly those with learning differences like dyslexia, ADD, low vision, and autism.