AI & Data Engineer, Data Discovery Services

Bristol Myers Squibb

Confirmed live yesterday High trust
Closes tomorrow

Quick summary

Work type
On-site
Location
Princeton, NJ
Salary
$87,810–$106,399 / yr
Employment
Full-time
Posted
12 days ago
Freshness
Confirmed live yesterday
Closes
Oct 5, 2026 (soon)

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $177k
This role $97k
$72k most similar roles pay here $236k

This role pays less than 99% of similar roles. Most pay $137,888–$215,162 — the shaded band above. At the midpoint, this role pays about $97k versus about $177k for comparable roles.

Based on 240 similar postings.

Employer

About Bristol Myers Squibb

Bristol Myers Squibb is a global biopharmaceutical company committed to discovering, developing and delivering innovative medicines to patients.

Bristol Myers Squibb currently has 37 open roles on FindRole.

Listed pay typically runs $130,000–$166,654 across 31 roles with salary data.

Most-posted roles

View all roles at Bristol Myers Squibb

At a glance

TL;DR · AI & Data Engineer, Data Discovery Services

As an AI & Data Engineer, Data Discovery Services, you will join the Enterprise Data Platforms team to build and maintain the data foundation powering search and discovery across the organization. You will serve as a hands-on Python developer creating pipelines for metadata enrichment, transformations, and platform integrations. Your work involves building a semantic knowledge layer using vector embeddings and chunking to support retrieval-augmented generation for large language models. The role requires proficiency in Python and SQL, with experience in Databricks, AWS Glue, OpenSearch, and various AWS services like S3 and Lambda. You will also develop API endpoints and Model Context Protocol servers while managing data quality across complex pipelines. This position solves the critical problem of making enterprise data findable and actionable through advanced search technologies, graph databases, and AI-powered discovery capabilities.

What you'll do

  • Develop Python pipelines to extract, enrich, and publish metadata from enterprise data catalogs.
  • Build and maintain semantic knowledge layers including document chunking and vector embedding generation.
  • Tune and optimize search indexes by adjusting analyzers, boosting fields, and testing queries.
  • Maintain integrations that sync ontology and taxonomy changes into the discovery platform.
  • Resolve data pipeline issues across Databricks and AWS Glue while implementing quality checks.
  • Build API endpoints and Model Context Protocol servers to expose search capabilities to AI agents.
  • Design metadata pipelines to map cross-domain dataset relationships with confidence scores.
  • Analyze search patterns and user feedback to improve the overall discovery experience.

What we're looking for

  • Bachelor's degree in computer science, Data Science, Information Science, Engineering, or a related field.
  • Master's degree (preferred).
  • Demonstrated proficiency in data engineering or software engineering with a track record of delivering production data pipelines.
  • Proficiency in Python and SQL.
  • Experience with Databricks and AWS Glue for data pipelines and transformations.
  • Familiarity with AWS cloud services, OpenSearch, or Elasticsearch.
  • Understanding of metadata management and data cataloging concepts.
  • Experience with semantic knowledge layers, RAG patterns, vector search, or graph databases (preferred).

More like this

Similar roles

Data Engineer

Booz Allen Hamilton

Albuquerque, NM 30 days ago $61,900–$141,000
ETL ELT Python PostgreSQL AWS Docker Kubernetes Terraform PySpark CI/CD RESTful APIs CloudFormation CDK Data Modeling Big Data infrastructure-as-code

Data Engineer, AI Enablement

AbbVie

Chicago, IL 18 days ago $84,500–$162,000
SQL Python ETL ELT Airflow Data Engineering Machine Learning RAG Knowledge Graphs Vector Databases Databricks Spark Snowflake Neo4j AWS cloud infrastructure Agile Jira GxP
5+ yrs exp

AI Data Platform Engineer

Apple Inc

Cupertino, CA 64 days ago $150,400–$277,600
Python SQL Java Scala Airflow Kubeflow MLflow Spark PySpark Kafka Ray Pandas Iceberg Delta Lake RAG Vector Databases Kubernetes Docker CI/CD AWS Azure GCP
5+ yrs exp

AI Data & Knowledge Engineer

Apple Inc

Cupertino, CA 18 days ago $149,200–$249,000
Python SQL R Java Spark Hadoop Snowflake Redshift Databricks AWS Azure Google Cloud Apache Airflow Git CI/CD Vector Databases Knowledge Graphs RAG ETL REST JSON Parquet
5+ yrs exp

Data and Applied AI Engineer

Apple Inc

Seattle, WA 42 days ago $142,300–$263,300
Java Scala Big Data Microservices LLMs AI Agents REST APIs Distributed Systems Data Pipelines Algorithms Data Structures