Co-Op Data Scientist

IBM

Confirmed live yesterday Low trust

Quick summary

Work type
On-site
Location
State College, PA
Employment
Intern
Posted
26 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $162k
$85k most similar roles pay here $232k

This listing doesn't post a salary. Most similar roles pay $126,800–$196,363.

Based on 240 similar postings.

Employer

About IBM

IBM is a US-based global technology company providing hybrid cloud, AI, consulting, enterprise software, and IT infrastructure products and services.

IBM currently has 506 open roles on FindRole.

Listed pay typically runs $174,632–$197,165 across 5 roles with salary data.

Most-posted roles

View all roles at IBM

At a glance

TL;DR · Co-Op Data Scientist

Co-Op Data Scientist 2027 works within IBM Consulting to support clients through hybrid cloud and AI initiatives. This role involves participating in mentored analytical support projects where you will apply statistical methods to predict behaviors, develop data integrations by writing programs to cleanse and model data, and build data pipelines as a tech-driven data transformer. You will also work on generative AI projects involving large datasets to train models for natural language processing and audio processing. The role requires proficiency in Python and familiarity with libraries like scikit-learn, SciPy, pandas, and PyTorch. Additional preferred skills include experience with SQL, Spark, cloud platforms like IBM Cloud or AWS, and specific tools such as LangChain and watsonx. You will communicate complex findings to both technical and non-technical audiences while solving problems in the consulting space.

What does a Data Scientist earn?

Median $162025 from 276 postings across 61 companies.

See salary data

What you'll do

  • Implement generative AI models and algorithms for conversational computing, natural language processing, and audio processing.
  • Train and evaluate generative models using large datasets while optimizing performance through various techniques.
  • Write efficient and reusable programs to cleanse, integrate, and model data for actionable insights.
  • Build data pipelines using Python to extract and transform data from repositories to consumers.
  • Utilize ETL tools and cloud platforms to manage and integrate complex data systems.
  • Communicate analytical results and complex findings to both technical and non-technical audiences.
  • Apply statistical methods and analytical rigor to predict behaviors for client projects.

What we're looking for

  • Must have a High School Diploma or GED.
  • Bachelor's Degree (preferred).
  • Currently pursuing a quantitative degree in Computer Science, Statistics, Mathematics, Engineering, IST, or a related field with an expected graduation of May 2026 or later.
  • Familiarity with one or more scripting languages like Python or a proven computer science foundation.
  • Ability to work on-site in State College, PA during the academic year and maintain specific hourly availability.
  • Must have the ability to obtain and maintain a Federal clearance if necessary for client assignments.
  • Experience with statistical analysis, data mining, databases (SQL), cloud platforms, or machine learning libraries (preferred).
  • Familiarity with NLP/LLM/GenAI tools such as AutoGen, Langgraph, Lanchain, watsonx orchestrate, or OpenAI (preferred).

More like this

Similar roles

Co-Op Data Engineer

IBM

University Park, PA 26 days ago
Machine Learning Generative AI Agentic AI Data Science SQL JavaScript HTML XML AWS Azure Microservices APIs Middleware Agile Copilot ChatGPT

Intern Data Scientist, AI & Data Analytics

IBM

26 days ago
Python SQL R Machine Learning GenAI RAG Agentic AI pandas NumPy SciPy scikit-learn PyTorch TensorFlow LangChain LangGraph LlamaIndex Cloud Platforms APIs GitHub Copilot Agile

Federal Data Engineering Co-Op

IBM

University Park, PA 25 days ago
Python SQL Spark Scikit-learn SciPy Pandas PyTorch IBM Cloud Azure AWS ETL NLP LLM GenAI data-engineering Statistical Analysis Data Mining Agile

Associate Data Scientist, AI & Data Analytics

IBM

Chicago, IL 26 days ago
Python R SQL TensorFlow GenAI Agentic AI Machine Learning Data Science RAG LangChain LangGraph LlamaIndex Semantic Kernel AutoGen AWS Azure Google Cloud Git CI/CD MLOps LLMOps Vector Databases APIs

Associate Data Scientist, AI & Data Analytics

IBM

26 days ago
Python R SQL TensorFlow Machine Learning Generative AI Agentic AI RAG LangChain LangGraph LlamaIndex Semantic Kernel AWS Azure Google Cloud Git CI/CD MLOps LLMOps Vector Databases Statistics

Associate Data Scientist, AI & Data Analytics

IBM

26 days ago
Python R SQL TensorFlow Machine Learning Generative AI Agentic AI RAG LangChain LangGraph LlamaIndex Semantic Kernel AutoGen MLOps LLMOps CI/CD Git AWS Azure Google Cloud Vector Databases