Data Engineer, Clinical Operations

Bristol Myers Squibb

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Princeton, NJ
Salary
$87,810–$106,399 / yr
Posted
31 days ago
Freshness
Confirmed live yesterday
Closes
Sep 28, 2026

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $169k
This role $97k
$73k most similar roles pay here $226k

This role pays less than 98% of similar roles. Most pay $126,800–$211,200 — the shaded band above. At the midpoint, this role pays about $97k versus about $169k for comparable roles.

Based on 240 similar postings.

Employer

About Bristol Myers Squibb

Bristol Myers Squibb is a global biopharmaceutical company committed to discovering, developing and delivering innovative medicines to patients.

Bristol Myers Squibb currently has 33 open roles on FindRole.

Listed pay typically runs $125,480–$152,049 across 33 roles with salary data.

Most-posted roles

View all roles at Bristol Myers Squibb

At a glance

TL;DR · Data Engineer, Clinical Operations

Data Engineer, Clinical Operations will join the Global Drug Development IT group to support the Cross Study Operations and Specimen Management domain. This individual contributor role focuses on building reliable, secure data ecosystems for the full data product lifecycle, including ingesting, storing, processing, governing, and interacting with complex clinical datasets. The successful candidate will design and maintain production-grade pipelines, develop semantic models, and implement automated solutions for specimen tracking and biobanking workflows. Key technical requirements include expertise in Databricks, Delta Lake, Unity Catalog, MLflow, and Mosaic AI, alongside proficiency in Python, SQL, Spark, and PySpark. Additionally, the role requires experience with Generative AI frameworks, including RAG, fine-tuning, and vector embeddings. The work addresses critical life sciences challenges by improving data interoperability, traceability, and insight generation across cross-study clinical research and development programs.

What does a Data Engineer earn?

Median $183138 from 214 postings across 58 companies.

See salary data

What you'll do

  • Design, build, and maintain scalable production-grade data pipelines for cross-study operations and specimen management.
  • Develop and enhance data solutions to ensure interoperability and scalability across clinical R&D programs.
  • Optimize platform components using Databricks Delta Lake, cloud-native parallel processing, caching, and partitioning.
  • Implement robust ETL/ELT pipelines and data models using semantic modeling for complex life sciences datasets.
  • Enforce data governance, manage metadata, and ensure end-to-end lineage using Databricks Unity Catalog.
  • Build and deploy GenAI and NLP-driven applications to improve specimen traceability and operational efficiency.
  • Develop and manage machine learning models at scale using Databricks Mosaic AI and MLflow.
  • Provide technical guidance and mentorship to junior team members on Databricks and engineering best practices.

What we're looking for

  • Must have at least 2 years of hands-on experience in Data Engineering, Analytics, and AI/ML in a cloud environment.
  • Requires hands-on expertise with Databricks tools including Delta Lake, Unity Catalog, Workflows, Mosaic AI, and MLflow.
  • Must possess strong proficiency in Python, SQL, Spark (including PySpark), and Generative AI frameworks.
  • Requires experience with LLM architectures, RAG, prompt engineering, and agentic frameworks.
  • Must have expertise in cloud-native data platforms, ETL/ELT pipeline design, data modeling, and semantic analytics for large datasets.
  • Experience delivering production-grade GenAI applications and predictive models for clinical or research functions is required.
  • A Databricks certification (e.g., Data Engineer Associate or Professional) is preferred.
  • Knowledge of life sciences R&D and clinical trial operations is highly preferred.

More like this

Similar roles

Senior Manager, Data Engineer, Clinical Operations

Bristol Myers Squibb

Princeton, NJ 37 days ago $153,770$186,327
Databricks Python SQL Spark PySpark GenAI LLM RAG Delta Lake Unity Catalog MLflow Mosaic AI ETL ELT Data Modeling Semantic Search Vector Embeddings Containerization DevOps Veeva Vault
7+ yrs exp

Data Engineer

Booz Allen Hamilton

Albuquerque, NM 8 days ago $61,900$141,000
ETL ELT Python PostgreSQL AWS Docker Kubernetes Terraform PySpark CI/CD RESTful APIs CloudFormation CDK Data Modeling Big Data infrastructure-as-code

Data Engineer

Apple Inc

San Diego, CA 153 days ago $137,500$250,700
Python Oracle SQL Cassandra Linux Git Data Pipelines MLOps Machine Learning Data Architecture SQL NoSQL Data Visualization Business Intelligence STDF Software Engineering Data Engineering
3+ yrs exp

Data Engineer

Booz Allen Hamilton

Huntsville, AL 22 days ago $77,500$176,000
ETL ELT Python SQL Scala Java Spark Databricks Hadoop Hive AWS Azure Google Cloud Kafka MongoDB Cassandra Redshift Snowflake UNIX Linux Shell Scripting Agile

Data Engineer

Booz Allen Hamilton

Washington, DC 30 days ago $77,600$176,000
Python R SQL GenAI ETL Data Mining Machine Learning Statistical Analysis Agile Data Architecture Data Engineering
8+ yrs exp

Data Engineer

Booz Allen Hamilton

Arlington, VA 16 days ago $62,000$141,000
Python ETL ELT Big Data AWS Redshift SageMaker Databricks Apache Spark NVIDIA CUDA Terraform CloudFormation SQL PaaS SaaS Agile