SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 140,000 - 200,000 / annual
Speechify is a text-to-speech platform used by over 50 million people globally to convert PDFs, books, documents, articles, and websites into audio. The company was named Chrome Extension of the Year by Google and received Apple's 2025 Design Award for Inclusivity. With ~200 employees distributed globally and no physical office, Speechify operates as a fully remote organization.
You'll join the Data side of the AI team, responsible for all aspects of data collection to support model training operations. The team builds high-quality datasets at petabyte scale through tight integration of infrastructure, engineering, and research. This is a hands-on engineering role with significant impact on Speechify's next-generation AI models.
Key responsibilities include: sourcing new audio data and integrating it into the ingestion pipeline; operating and extending cloud infrastructure (GCP, Terraform) for data ingestion; collaborating with research scientists to optimize the cost/throughput/quality frontier; and working with leadership to define the AI team's dataset roadmap for consumer and enterprise products.
You should have a BS/MS/PhD in Computer Science or related field, 5+ years of software development experience, proficiency with bash/Python in Linux, Docker, Infrastructure-as-Code, and experience with at least one major cloud provider (GCP preferred). Experience with web crawlers and large-scale data processing is a plus. The role requires strong communication, adaptability, and comfort with scrappy problem-solving in a fast-moving environment.
Speechify offers competitive compensation, equity, a hands-off management approach, and the opportunity to impact a product used by millions, including people with learning differences like dyslexia, ADD, low vision, and autism.