$87,810 – $106,399
Listed on Bristol Myers Squibb’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role supports Bristol Myers Squibb's clinical data infrastructure by designing and operating scalable data pipelines, ETL workflows, and GenAI applications for cross-study operations and specimen management. It suits experienced data engineers who combine cloud platform expertise with machine learning knowledge and want to apply their skills to life sciences and clinical trial workflows.
Our summary, not Bristol Myers Squibb’s wording. The full posting is on their site.
Skills this role names
- Cloud Platforms
- Data Modeling
- Databricks
- DevOps
- GenAI (Generative AI)
- MLflow
- Prompt Engineering
- PySpark
- Python
- Retrieval-Augmented Generation (RAG)
- SPARK
- SQL
Log in to see which of these are already on your profile.
What they ask for
Required
- 5+ years of hands-on data engineering, analytics, or AI/ML experience
- Expertise with Databricks including Delta Lake, Unity Catalog, and Workflows
- Cloud-native data platform and ETL/ELT pipeline design experience
- Proficiency in Python, SQL, and PySpark
- Experience delivering production-grade GenAI applications or predictive models
- Strong stakeholder engagement and communication skills
Nice to have
- Databricks certification (Data Engineer Associate or Professional)
- Experience with life sciences, clinical trial operations, or specimen/biobanking workflows
- Functional knowledge of Life Sciences R&D and clinical trial operations
- Hands-on experience with LLM architectures, RAG, and agentic frameworks