Remote job
Data Engineering Intern
Job details
About this role
Role overview
This is a four-month, full-time paid data engineering internship designed to give hands-on professional experience on applied AI and data projects. The intern works alongside senior specialists and engineers in a fully remote, globally distributed team that ships work for clients across many industries. The role blends data pipeline work, exploration and visualization, and direct exposure to client communications.
Responsibilities
- Help design, build, and maintain data pipelines that deliver clean, reliable, and timely data. - Contribute to the implementation and optimization of ETL pipelines. - Integrate data from multiple sources into warehouses, data lakes, and lakehouses. - Support data management tasks including cleaning, validation, and transformation. - Help translate business objectives into data models and metrics that track progress. - Participate in client communications by gathering requirements and communicating deliverables. - Explore and visualize datasets, flagging quality problems and distribution differences that could affect downstream performance. - Identify data quality issues and propose improvements to the codebase.
Requirements
- Proficiency in English; the CV must be submitted in English, and all meetings and client work happen in English. - Basic knowledge of Python and common data libraries. - Basic knowledge of SQL and database engines. - Experience manipulating and visualizing datasets. - Strong problem-solving skills. - Bonus: AWS knowledge, (Py)Spark, Airflow, data lakes, data warehouses, or Git.
Nice to have
- A curious mindset that wants to learn across different industries with a modern tech stack. - Comfort working autonomously and positively in a fully remote, distributed team. - Collaborative, adaptable approach with a startup pace and dependable delivery.
Benefits and work setup
- Fully remote and flexible. - Every other Friday off. - Paid sick days and local holidays. - Fitness subscription.