Data Engineering Intern

Full-time
Portugal
Entry Level
Posted 2 hours ago
Apply for this position → Go ad-free with Premium ×

Duration: Three months Format: Full time (40 hrs/week), paid

In the last year at Loka, our teams launched almost 200 GenAI projects for companies of all kinds, including the world's number 1 GenAI reading tutor, a startup that transforms homes into batteries and a leading cancer fighting laboratory. To cap it off, at the end of 2024 Loka was recognized by AWS as Innovation Partner of the Year, outshining 150,000 partners for the title. And we did it all while enjoying every other Friday off 😎.

As a Data Engineering Intern, you'll gain hands on professional experience supporting Loka's certified specialists, technical experts and PhDs, all while elevating your skillset, building a portfolio and launching projects you're proud of.

What You'll Do

  • Assist in designing, developing and maintaining data pipelines to ensure clean, reliable and timely data.

  • Collaborate with the team to implement and optimize ETL pipelines.

  • Integrate data from various sources into warehouses, data lakes and lakehouses.

  • Support data management tasks, including data cleaning, validation and transformation.

  • Understand business objectives and help develop data models that support them, along with metrics to track progress.

  • Participate in client communications by helping gather requirements and communicate deliverables.

  • Explore and visualize data with a careful eye for issues that require cleaning, as well as differences in data distribution that may affect performance after deployment.

  • Identify data quality issues and opportunities to improve the codebase.

What You'll Bring

Experience & Technical Skills

  • Proficient in English.

  • Basic knowledge of Python and data libraries.

  • Basic knowledge of SQL and database engines.

  • Experience visualizing and manipulating datasets.

  • Strong problem solving skills.

  • Bonus: AWS knowledge, (Py)Spark, Airflow, data lakes and data warehouses, git.

Additional Requirements

  • Excellent English, we work entirely in English for meetings, client calls and business communications.

  • CV submitted in English.

Personality Profile

  • Curious: You're ambitious to learn and grow in different industries using a modern tech stack.

  • Autonomous and positive: You excel in a fully remote, globally distributed team.

  • Team player: You enjoy a collaborative approach.

  • Adaptable: You operate with a startup mindset and move at a startup pace.

  • Dependable: You can be trusted to deliver high quality work.

Benefits

  • Every other Friday off

  • Remote and flexible

  • Paid sick days and local holidays

  • Fitness subscription

Your achievements matter to us! Ensure your CV and LinkedIn profile are up to date and accurately reflect your experience.

Go ad-free with Premium ×
Apply for this position →
About the Job
Full-time
Portugal
Entry Level
Posted 2 hours ago
Check if your resume is a good fit
25/100
Get Full Report
+ 1,284 new jobs added today
30,000+
Remote Jobs

Don't miss out — new listings every hour

Join Premium

Data Engineering Intern

Duration: Three months Format: Full time (40 hrs/week), paid

In the last year at Loka, our teams launched almost 200 GenAI projects for companies of all kinds, including the world's number 1 GenAI reading tutor, a startup that transforms homes into batteries and a leading cancer fighting laboratory. To cap it off, at the end of 2024 Loka was recognized by AWS as Innovation Partner of the Year, outshining 150,000 partners for the title. And we did it all while enjoying every other Friday off 😎.

As a Data Engineering Intern, you'll gain hands on professional experience supporting Loka's certified specialists, technical experts and PhDs, all while elevating your skillset, building a portfolio and launching projects you're proud of.

What You'll Do

  • Assist in designing, developing and maintaining data pipelines to ensure clean, reliable and timely data.

  • Collaborate with the team to implement and optimize ETL pipelines.

  • Integrate data from various sources into warehouses, data lakes and lakehouses.

  • Support data management tasks, including data cleaning, validation and transformation.

  • Understand business objectives and help develop data models that support them, along with metrics to track progress.

  • Participate in client communications by helping gather requirements and communicate deliverables.

  • Explore and visualize data with a careful eye for issues that require cleaning, as well as differences in data distribution that may affect performance after deployment.

  • Identify data quality issues and opportunities to improve the codebase.

What You'll Bring

Experience & Technical Skills

  • Proficient in English.

  • Basic knowledge of Python and data libraries.

  • Basic knowledge of SQL and database engines.

  • Experience visualizing and manipulating datasets.

  • Strong problem solving skills.

  • Bonus: AWS knowledge, (Py)Spark, Airflow, data lakes and data warehouses, git.

Additional Requirements

  • Excellent English, we work entirely in English for meetings, client calls and business communications.

  • CV submitted in English.

Personality Profile

  • Curious: You're ambitious to learn and grow in different industries using a modern tech stack.

  • Autonomous and positive: You excel in a fully remote, globally distributed team.

  • Team player: You enjoy a collaborative approach.

  • Adaptable: You operate with a startup mindset and move at a startup pace.

  • Dependable: You can be trusted to deliver high quality work.

Benefits

  • Every other Friday off

  • Remote and flexible

  • Paid sick days and local holidays

  • Fitness subscription

Your achievements matter to us! Ensure your CV and LinkedIn profile are up to date and accurately reflect your experience.