Remote job
Backend Engineer, Data Infrastructure
Job details
About this role
Role overview Build and maintain the backend services and data pipelines that collect, process, and serve multilingual digital discourse data. You will help ensure reliable data quality and turn analytical needs into production systems that support measurement and reporting.
Responsibilities - Develop and operate Python backend services and data pipelines on Google Cloud Platform. - Write and optimize SQL queries for high-volume data. - Validate and deduplicate incoming data, and monitor quality throughout processing. - Translate data science requirements into dependable production services. - Add logging and metrics so pipeline problems are visible early. - Take part in code reviews and help maintain engineering standards.
Requirements - 2–3 years of professional backend or data engineering experience. - Strong working knowledge of Python and SQL. - Hands-on experience with GCP services such as BigQuery, Cloud Functions, or Pub/Sub. - Ability to break down ambiguous data problems into clear, testable steps. - Familiarity with Git and standard continuous integration practices. - Clear written communication when working with data science and product colleagues.
Nice to have - Experience with streaming or near-real-time pipelines. - Exposure to NLP or text-processing systems, especially for non-English languages. - Familiarity with Docker, Cloud Run, or Kubernetes.
Benefits and work setup Full-time, remote-first position with a regional-market competitive salary. The team is small and senior, with low process overhead and ownership of infrastructure used by regional institutions.