707,945open jobs
41,960companies
101,748added this week
Browse all
Salary
$88k – $106k per year
Location
In office (Princeton)
Employment
Full-Time
Overview
Company
Impact
Profile match
Bristol Myers Squibb is an American pharmaceutical company whose predecessors date to 1858 and 1887 and which took its present form in a 1989 merger. Its portfolio is concentrated in oncology, haematology, immunology and cardiovascular disease, including the checkpoint inhibitor Opdivo, the blood thinner Eliquis co-marketed with Pfizer, the cell therapies acquired with Celgene, and newer products such as the schizophrenia treatment Cobenfy. Headquartered in Princeton, New Jersey and listed on the New York Stock Exchange, it faces a heavy sequence of patent expiries and has been buying and licensing assets to replace them.

At Bristol Myers Squibb, our employees often ask, “Who are you working for?”-a question that fuels collaboration, accountability, and urgency in our work. Our purpose-driven culture inspires us to discover, develop, and deliver innovative medicines to prevail over serious diseases. We offer uniquely interesting and meaningful work, opportunities for growth, and a supportive environment that values inclusion, wellbeing, flexibility, and comprehensive benefits. This is work that transforms the lives of patients, and the careers of those who do it.

Position Summary:

Join the Data Discovery Services team within Enterprise Data Platforms, where we deliver and maintain the data foundation platforms that power discovery, search, and data accessibility across the Bristol Myers Squibb enterprise. We run multiple search and discovery services used across the company -- and we're building the next generation of AI-powered discovery on top of them. Our work makes enterprise data findable, accessible, and actionable for teams across the organization. At the core of this is a semantic knowledge layer -- metadata, taxonomies, and relationships that describe what data means and how it connects -- curated as a data inventory that helps AI work reliably across the enterprise. This is a high-impact team where engineering, search, and applied AI come together to solve real problems at scale.

As an AI / Data Engineer, you'll be a hands-on Python developer building the pipelines and integrations that make enterprise data more discoverable. Your primary focus is data engineering -- pipelines, metadata enrichment, transformations, and platform integrations. You'll also contribute to search and AI-powered retrieval as you grow into the role. Working alongside data engineers, search engineers, and data scientists, this is a hands-on engineering role -- you'll write code, build pipelines, ship features, and own what you deliver.

Why Join Us?

  • Work with a modern stack -- Databricks, Amazon Web Services (AWS), OpenSearch, vector search, semantic knowledge layers, graph databases, and AI agents.

  • Build real AI-powered discovery capabilities, not proofs of concept.

  • Grow your skills across data engineering, search, and applied AI on the same team.

  • Use AI-assisted development tools (Claude, Copilot) in your daily workflow.

  • Contribute to open-source projects and shared accelerators.

  • Clear path to grow into senior engineering, search specialization, or AI engineering roles.

  • Make enterprise data findable and accessible for teams working to improve patient outcomes.

Job Responsibilities:

As an AI / Data Engineer, you'll be a hands-on Python developer building the pipelines and integrations that make enterprise data more discoverable. Your primary focus is data engineering - pipelines, metadata enrichment, transformations, and platform integrations.

You'll also contribute to search and AI-powered retrieval as you grow into the role. Working alongside data engineers, search engineers, and data scientists, this is hands-on engineering role - you'll write code, build pipelines, ship features, and own what you deliver.

  • Build and maintain Python pipelines that pull metadata from enterprise data catalogs, enrich it with taxonomy tags and ownership information, and publish it to the discovery platform.

  • Tune and optimize search indexes -- adjust analyzers, boost fields, and test queries -- to ensure results match what users need.

  • Build a semantic knowledge layer -- chunking documents, generating vector embeddings, and enriching them with semantic knowledge metadata -- to grow a data inventory that supports retrieval-augmented generation (RAG) and helps AI systems and large language models (LLMs) find and use the right context.

  • Maintain integrations that sync ontology and taxonomy changes into the discovery platform, so classifications stay current.

  • Investigate and resolve data pipeline issues across Databricks and AWS Glue, trace root causes through metadata enrichment flows, and add data quality checks to prevent recurrence.

  • Build API endpoints and Model Context Protocol (MCP) servers that expose search and metadata capabilities to applications and AI agents.

  • Design metadata pipelines that map cross-domain dataset relationships and add them to the cross-domain join catalog with confidence scores.

  • Analyze search patterns, capture user feedback, and improve the discovery experience so the system learns and improves over time.

Qualifications & Experience:

Required

  • Bachelor’s degree in computer science, Data Science, Information Science, Engineering, or a related field. Master's degree preferred.

  • Demonstrated proficiency in data engineering, software engineering, or a related technical discipline, with a track record of delivering production data pipelines.

  • Proficient Python skills -- this is your primary language day-to-day.

  • Proficiency in Structured Query Language (SQL).

  • Experience with Databricks and AWS Glue for data pipelines and transformations.

  • Solid data engineering fundamentals: extract-transform-load (ETL/ELT) patterns, data modeling, data quality, and pipeline orchestration.

  • Familiarity with AWS cloud services (S3, Lambda, API Gateway, Glue).

  • Experience with OpenSearch or Elasticsearch.

  • Understanding of metadata management and data cataloging concepts.

  • Effective problem-solving skills and willingness to learn.

  • Good communication skills and ability to work collaboratively in a team.

Preferred Qualifications:

  • Experience with semantic knowledge layers, RAG patterns, vector search technologies, and building AI-ready data inventories.

  • Familiarity with semantic search, embeddings, chunking strategies, relevance tuning, and semantic knowledge metadata (entity relationships, taxonomies, context enrichment).

  • Exposure to AI agent patterns, MCP, or large language model orchestration frameworks.

  • Experience with ontology or taxonomy technologies (such as Turtle, Resource Description Framework, Web Ontology Language, or SPARQL query language) or management platforms.

  • Familiarity with graph databases or knowledge graph technologies.

  • Experience with metadata enrichment, data lineage, or data quality frameworks.

  • Exposure to Azure OpenAI, Google Vertex AI, or Amazon Bedrock.

  • Experience with Docker, Elastic Container Service (ECS), or CloudFormation.

  • Prior exposure to pharma or life sciences.

We hire for skills and capabilities, not just credentials - if this role excites you, but doesn’t perfectly match your resume, we encourage you to apply anyway.

Compensation Overview:

Princeton - NJ - US: $87,810 - $106,399

The starting pay range(s) listed above is for full-time employees (FTE). You may also be eligible for additional discretionary incentive cash and stock opportunities. We determine starting pay thoughtfully - carefully considering the nature of the role, required skills, work location, schedule and the knowledge and experience you bring. Final compensation is guided by pay equity principles and applicable employment laws. Compensation programs are reviewed on an ongoing basis and may be adjusted over time to reflect evolving market factors, and individual, team or Company performance.

Benefits:

Subject to the terms and conditions of the applicable plans then in effect, you may be eligible to participate in our comprehensive benefit plans - including wellbeing support, retirement and financial protection benefits, and insurance offerings (medical, dental, vision, life and disability).

U.S.-based exempt employees are eligible for Flexible Time Off (FTO), which provides paid time off without a set accrual limit, subject to manager approval, along with 11 paid company holidays each year.

Non-exempt employees, RayzeBio employees, and employees located in Puerto Rico receive 160 hours of paid vacation annually for new hires (subject to manager approval), 11 paid company holidays, and 3 optional holidays.

Depending on eligibility, employees may also have access to additional time-off benefits, including paid sick leave, up to two paid volunteer days per year, summer hours flexibility, and leaves of absence for medical, personal, parental, caregiver, bereavement, or military needs. Eligible employees also enjoy an annual Global Shutdown between Christmas Day and New Year's Day.

U.S.-based job seekers can explore full benefit offerings athttps://careers.bms.com/benefits

How We Work

Where you work matters - because collaboration, innovation and patient impact happen in many settings. Our roles are structured across four work models: site-essential, site-by-design, field-based and remote-by-design. The model assigned to this role is based on its core responsibilities. Learn more at https://careers.bms.com/ways-of-working.

Supporting People with Disabilities

BMS is dedicated to ensuring that people with disabilities can excel through a transparent recruitment process, reasonable workplace accommodations/adjustments and ongoing support in their roles. Applicants can request a reasonable workplace accommodation/adjustment prior to accepting a job offer. If you require reasonable accommodations/adjustments in completing this application, or in any part of the recruitment process, direct your inquiries to [email protected]. Visit careers.bms.com/eeo-accessibility to access our complete Equal Employment Opportunity statement.

Candidate Rights

BMS will consider qualified applicants with arrest and conviction records, pursuant to applicable laws in your area.

For roles based in Los Angeles County only: If you live in or expect to work from Los Angeles County if hired for this position, please visit this page for important additional information: https://careers.bms.com/california-residents/

Data Protection

We will never request payments, financial information, or social security numbers during our application or recruitment process. Learn more about protecting yourself at https://careers.bms.com/fraud-protection.

Any data processed in connection with role applications will be treated in accordance with applicable data privacy policies and regulations.

If this posting is missing required information required by local law or incorrect, contact BMS at [email protected] with the Job Title and Requisition number. Do not send application-related inquiries to this email. To check your application status, please login to your Candidate Home Account.

R1604590 : AI & Data Engineer, Data Discovery Services

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
707,945 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Princeton
$17k – $43k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Pune
Python
SQL
Python
pySpark
Databases
Databricks
Azure Cosmos DB
AI/ML
LangChain
Spark
MLFlow
XGBoost
Scikit-learn
Prompt Engineering
NLP
LightGBM
spaCy
TensorFlow
Pandas
NumPy
PyTorch
LLM
RAG
NLTK
OpenAI
Hugging Face
Edge AI
Machine Learning
DevOps
Azure DevOps
GitHub Actions
Azure
CI/CD
Git
Docker
Kubernetes
Analytics
Power BI
Seaborn
Matplotlib
Plotly
ETL/ELT
Azure Data Factory
Management
Agile
Scrum
Apply
$29k – $60k per year (Estimated) • Remote/Hybrid • 3+ years exp • Moscow
Python
SQL
Python
pySpark
Databases
PostgreSQL
Apache Iceberg
Apache Kafka
Trino
AI/ML
Hadoop
Spark
MLlib
DevOps
Kubernetes
Amazon S3
Analytics
ETL/ELT
Apply
Equity • Remote • Internship • Bachelor's Degree • United States
Python
JavaScript
TypeScript
SQL
C#
Node JS
Node JS
Nest.JS
Databases
ElasticSearch
Frontend
React.js
DevOps
AWS
Kubernetes
GitHub
Management
Slack
Confluence
Jira
Apply
Research Technician I 6 hours ago
$40k per year • In office • Part-Time • 2+ years exp • High School Diploma • College Station
Python
Apply
$52k – $144k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
SQL
Analytics
Power BI
Apply
$298k – $361k per year • In office • Full-Time • 10+ years exp • PhD • Princeton • Madison
Management
Agile
Apply
$86k – $104k per year • In office • Full-Time • 1+ year exp • Bachelor's Degree • New Brunswick
Apply
$178k – $216k per year • In office • Full-Time • 5+ years exp • PhD • Princeton
Apply
$68k – $84k per year • In office • Full-Time • 5+ years exp • High School Diploma • Devens
Management
Outlook
Microsoft Office
Apply
$58k – $115k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Devens
Cybersecurity
CAPA
Apply
$170k – $272k per year • In office • Full-Time • 15+ years exp • PhD • Princeton
Python
C#
C++
C#
.NET
AI/ML
Computer Vision
Machine Learning
DevOps
CI/CD
Git
Linux
Windows
Management
Agile
Apply
$110k – $189k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Boston • New York • Princeton • Austin • Burlington
Apply
$70k – $119k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • Princeton
Python
Java
Databases
Apache Kafka
DevOps
Azure
AWS
Kubernetes
Management
Agile
Scrum
Apply
$159k – $193k per year • In office • Full-Time • PhD • Princeton • Cambridge • Brisbane
Apply
$28k – $49k per year (Estimated) • In office • 1+ year exp • Princeton
Apply
See all jobs
This is one of many
707,945 more open roles from verified company boards, updated every day.