Remote job
Software Engineer, Data Infrastructure & Acquisition - Kirkland, WA, USA
Job details
About this role
Role overview Build and run the data infrastructure that feeds large-scale model training for a text-to-speech reading product used by millions. The role covers sourcing audio data, operating the ingestion cloud stack, and working alongside scientists to keep improving cost, throughput, and quality. You will be part of a fully remote engineering organization, with Kirkland, Washington as the target location.
Responsibilities - Identify new sources of audio data and route them into the ingestion pipeline - Maintain and expand the cloud infrastructure that backs ingestion, currently running on Google Cloud Platform and provisioned with Terraform - Collaborate with research scientists to improve the cost, throughput, and quality frontier of dataset production - Co-author the team's dataset roadmap with AI leadership - Help power next-generation consumer and enterprise products with richer, larger-scale training data
Requirements - BS, MS, or PhD in Computer Science or a closely related field - 5+ years of professional software development experience - Proficiency in bash and Python scripting on Linux - Practical experience with Docker and Infrastructure-as-Code, plus professional work at a major cloud provider - Ability to juggle multiple tasks and respond to shifting needs - Strong written and verbal communication
Nice to have - Experience building web crawlers or running large-scale data processing pipelines
Benefits and work setup - 100% distributed, asynchronous team with no central office - Entrepreneurial group that values initiative, risk-taking, and hustle - Light-touch management meant to protect focused, independent work - Competitive compensation with bonus and equity; US base salary range of $140,000–$200,000 depending on experience - Product work that directly benefits users with dyslexia, ADD, low vision, concussions, autism, and other learning differences