Job details
About this role
Role overview Contribute to the data foundation of a widely used text-to-speech platform with more than 50 million users globally. The position focuses on acquiring and ingesting audio data at petabyte scale to train next-generation voice models, combining infrastructure engineering with research collaboration. Work happens in a fully distributed, asynchronous setup alongside scientists and engineers from leading tech companies.
Responsibilities - Discover and integrate new audio data sources into a high-throughput ingestion pipeline. - Run and extend cloud infrastructure on GCP, managed with Terraform, supporting continuous data flows. - Work side-by-side with research scientists to improve the cost, throughput, and quality balance of training datasets. - Help shape the dataset roadmap in partnership with AI leadership for future consumer and enterprise offerings. - Develop tooling and automation for large-scale data workflows in Linux environments.
Requirements - BS, MS, or PhD in Computer Science or a related technical field. - 5+ years of professional software development experience. - Proficiency scripting in bash and Python on Linux platforms. - Hands-on experience with Docker and Infrastructure-as-Code concepts. - Production experience with at least one major cloud provider, with GCP familiarity preferred. - Strong communication skills and adaptability to shifting priorities.
Nice to have - Background with web crawlers or large-scale data processing pipelines.
Benefits and work setup - Fully remote, asynchronous team with no central office. - Entrepreneurial culture that supports initiative, risk-taking, and ownership. - Competitive compensation and a friendly working atmosphere. - Opportunity to build products that directly help users with learning differences including dyslexia, ADHD, low vision, concussions, and autism.