Senior Specialist, Data Engineering - Digital Sciences, Analytical Research and Development
MSDAbout the role
Job Description
Our Scientists are our Inventors. We deliver robust scientific data and interpretation that drives the commercialization of breakthrough therapeutics using innovative thinking, state-of-the-art facilities and modern equipment. Our ability to excel depends on the integrity, knowledge, imagination, skill, diversity and teamwork of people like you. To this end, we strive to create an environment of mutual respect, encouragement and teamwork. As part of our global team, you’ll have the opportunity to collaborate with talented and dedicated colleagues while developing your career.
The Analytical Research and Development organization in Rahway, NJ is seeking a motivated Senior Specialist with technical expertise in data engineering for a role that primarily involves envisioning, building, and implementing digital solutions to support our company's pipeline of small molecule drugs, peptides, biologics, and vaccines. The successful candidate is expected to collaboratively work with synthetic, computational and analytical scientists, understand their problem statements, and collaboratively design and provide solutions. The candidate must be able to work proactively and independently and influence decisions and solutions spanning a wide diversity of problem statements across our company's Research Laboratories. The role will require the candidate to network across a range of departments and effectively collaborate with stakeholders. The candidate should exemplify positive ways of working which support diversity, inclusion, and a positive culture.
Establishing data workflows, data visualizations, and predictive tools to enable more effective identification, characterization, and development of novel medicines and vaccines is a key objective for our company. This position sits within the Digital Sciences team in the Analytical Enabling Capabilities sub-department of Analytical Research & Development. You will be part of a team working collaboratively across a wide range of areas impacting all aspects of the drug discovery and development pipeline. A diverse array of projects spanning data ingestion and analysis workflows to instrument metrology to predictive sciences will be encountered in this role. The core Digital Sciences team works with a networked group of digital champions across AR&D and has close connectivity to other digital facing teams across our company's Research Laboratories including critical IT collaborators.
Primary Responsibilities:
Design and development of data workflows/data pipelines in Python.
Meet with business clients/SMEs to gather requirements.
Working with IT to implement data workflows.
Manage project and timelines.
Estimation of duration of work.
Presentation of updates to collaborators.
Education Requirements:
Bachelor's Degree in Computer Science, Software Engineering, Computer Engineering, Electrical Engineering or Chemistry with strong programming capabilities.
Required Experience and Skills:
Cloud Services - AWS (MH: Lambda Function, Redshift, S3, DataSync)
Development of ETL Processes / Data Workflows / Data Pipelines / Data Wrangling / Data Ingestion.
Python 3.9+ software development
Python packages (pyodbc, pandas, boto3)
Python virtual environments - Anaconda and conda
IDEs - Visual Studio Code or PyCharm
Software design, development, and testing - unit testing and system testing
Version control - Git, GitHub
Databases - relational databases, SQL, data modeling and design
File Formats (XLXS, YAML, JSON, CSV, TSV)
Excellent verbal and written communications skills.
Work independently and be able to collaborate as a team.
Strive for continuous improvement and suggest innovative solutions to scientists’ common challenges related to data workflows.
Insatiable curiosity to learn with the end goal of understanding the scientific problem statement and designing programs, software, or solutions to meet the end users’ needs
Excellent verbal and written communications skills
Preferred Experience and Skills:
Cloud Services – AWS (Glue, Athena)
Development of ETL Processes / Data Workflows / Data Pipelines / Data Wrangling / Data Ingestion.
Python packages (cerberus, yaml, logging, openpyxl, numpy)
Python linters and type hints; regular expressions
Experience with data pipeline tools such as Dataiku or Databricks
TetraScience (JSON, IDS JSON, ASM, ADF)
Track record of impact in the
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s