Data Engineer Intern

IBM

Confirmed live yesterday Low trust

Quick summary

Work type
On-site
Location
Employment
Intern
Posted
26 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $151k
$93k most similar roles pay here $210k

This listing doesn't post a salary. Most similar roles pay $104,000–$198,450.

Based on 240 similar postings.

Employer

About IBM

IBM is a US-based global technology company providing hybrid cloud, AI, consulting, enterprise software, and IT infrastructure products and services.

IBM currently has 506 open roles on FindRole.

Listed pay typically runs $174,632–$197,165 across 5 roles with salary data.

Most-posted roles

View all roles at IBM

At a glance

TL;DR · Data Engineer Intern

The Data Engineer Intern 2027 joins a team focused on delivering consulting services and technology transformation across diverse industries. In this role, you will work on client projects to build data pipelines, cleanse and integrate data, and model information to provide actionable insights. You will also communicate complex findings to both technical and non-technical audiences. The position requires proficiency in programming languages such as Python, SQL, and JavaScript, along with familiarity with data engineering tools like Spark. Candidates should have a foundation in cloud platforms including IBM Cloud, AWS, or Azure, and may utilize technologies involving NLP, LLM, or GenAI. You will apply statistical methods to predict behaviors while navigating an agile environment. The work involves solving complex data integration challenges and transforming raw data into valuable assets for various business clients.

What does a Data Engineer earn?

Median $183138 from 214 postings across 58 companies.

See salary data

What you'll do

  • Contribute to real client projects across diverse industries to build a professional portfolio.
  • Apply analytical rigor and statistical methods to predict behaviors based on data.
  • Write efficient, reusable programs to cleanse, integrate, and model complex data sets.
  • Build data pipelines using Python to extract and transform data from various repositories.
  • Utilize ETL tools and cloud platforms to manage and move data for consumers.
  • Communicate technical findings and analytical results to both technical and non-technical audiences.
  • Evaluate model results to provide actionable, data-driven insights for clients.

What we're looking for

  • High School Diploma or GED required.
  • Bachelor's Degree (preferred).
  • Currently pursuing a quantitative degree in Computer Science, Statistics, Mathematics, Engineering, or a related field.
  • Strong interpersonal skills for collaboration and managing dynamic workloads in an agile environment.
  • Demonstrated leadership experience and ability to communicate effectively through active listening.
  • Familiarity with one or more scripting languages (Python preferred) or a proven computer science foundation.
  • Willingness to travel as needed.
  • Experience with statistical analysis, data mining, databases, cloud platforms, or AI concepts (preferred).

More like this

Similar roles

Data Engineer Intern

IBM

26 days ago
Python SQL Spark IBM Cloud AWS Azure Snowflake Databricks ETL NLP LLM GenAI Hybrid Cloud Data Engineering JavaScript Statistical Analysis Data Mining Agile

Data Engineer Intern

IBM

26 days ago
Python SQL Spark IBM Cloud AWS Azure Snowflake Databricks ETL NLP LLM GenAI Hybrid Cloud JavaScript Data Engineering Statistical Analysis Data Mining Agile

Intern Data Engineer, AI & Data Analytics

IBM

26 days ago
Python SQL Java Scala JavaScript Spark Kafka Airflow dbt Databricks Snowflake Delta Lake Linux RAG LangChain LangGraph GitHub Copilot ETL/ELT APIs Vector Databases

Intern Data Scientist, AI & Data Analytics

IBM

26 days ago
Python SQL R Machine Learning GenAI RAG Agentic AI pandas NumPy SciPy scikit-learn PyTorch TensorFlow LangChain LangGraph LlamaIndex GitHub Copilot Cloud Platforms APIs Data Mining

Intern Data Scientist, AI & Data Analytics

IBM

26 days ago
Python SQL R Machine Learning GenAI RAG Agentic AI pandas NumPy SciPy scikit-learn PyTorch TensorFlow LangChain LangGraph LlamaIndex Cloud Platforms APIs GitHub Copilot Agile

Intern Data Scientist, AI & Data Analytics

IBM

26 days ago
Python SQL R Machine Learning GenAI RAG Agentic AI Pandas NumPy SciPy scikit-learn PyTorch TensorFlow LangChain LangGraph LlamaIndex Cloud Platforms APIs GitHub Copilot Agile