Job details
About this role
Role overview This role focuses on the data acquisition and infrastructure side of an AI team that powers large-scale text-to-speech products used by millions of people worldwide. The position blends infrastructure engineering, scripting, and research collaboration to build petabyte-scale training datasets at low cost. It is a fully remote software engineering position based in or around Hartford, CT.
Responsibilities - Identify and onboard new sources of audio data into an ingestion pipeline. - Operate and extend the cloud infrastructure supporting ingestion, currently built on GCP and managed with Terraform. - Partner with research scientists to push the cost, throughput, and quality frontier for training data. - Help shape the dataset roadmap that feeds next-generation consumer and enterprise products. - Maintain reliable, scalable pipelines that integrate tightly with model training operations.
Requirements - BS, MS, or PhD in Computer Science or a related field. - 5+ years of industry software development experience. - Strong proficiency in bash and Python scripting on Linux systems. - Hands-on experience with Docker and Infrastructure-as-Code patterns. - Professional experience with at least one major cloud provider (GCP preferred). - Strong written and verbal communication skills, with the ability to juggle shifting priorities.
Nice to have - Experience designing web crawlers and large-scale data processing workflows.
Benefits and work setup - Fully remote, distributed work environment with no central office. - Hands-off management style and an asynchronous, entrepreneurial culture. - Competitive base salary, bonus, and equity (US-based range of $140,000–$200,000 depending on experience). - Mission-driven product used by people with dyslexia, ADHD, low vision, and other learning differences. - Work at the intersection of AI and audio in a fast-growing sector.