Data Engineer - Computational Biology
PfizerAbout the role
ROLE SUMMARY
You will apply strong bioinformatics and cloud engineering practices to develop, operationalize and evolve production bioinformatics pipelines to deliver reliable data products for our Pfizer R&D research units.
You will be a key contributor to a dynamic and pioneering team dedicated to advancing a cutting‑edge *omics ecosystem platform for Pfizer R&D. You will leverage your expertise to design innovative approaches that extract valuable insights from Pfizer’s proprietary and external datasets, enabling the generation of testable hypotheses across the entire drug discovery value chain.
ROLE RESPOSIBILITIES
Developing, deploying, and operating production‑grade Nextflow pipelines on cloud infrastructure, ensuring scalable and reproducible execution.
Owning pipeline lifecycle management, including upgrades, troubleshooting, performance tuning, and reliability improvements for reusable workflows.
Implementing DevOps best practices for pipelines and platform services (e.g., CI/CD, automation, and engineering tooling).
Partnering with wet‑lab and research scientists to translate data analysis requirements into robust, production‑ready pipeline and platform solutions.
Developing and evolving an omics data platform that enables efficient, scalable processing and delivery of *omics datasets as reliable data products.
Driving collaborations with external partners and vendors to strengthen pipeline quality, sustainability, and adoption of best practices.
BASIC QUALIFICATIONS
PhD in Computational Biology, Biology, Physics, Statistics, or a related technical discipline
Masters in Computational Biology, Biology, Physics, Statistics, or a related technical discipline and a minimum of two years of experience developing data products and data integration solutions in a research or industry environment
Single‑cell/NGS, functional genomics, genetics, or proteomics data analysis experience
Hands‑on experience developing Nextflow pipelines for processing NGS data
Strong full‑stack programming skills with a focus on Python
Experience solving complex analyses/problems in a timely fashion
Excellent communication and collaboration skills with experience working effectively in cross-functional teams
PREFERRED QUALIFICATIONS
Background or demonstrated interest in life sciences, pharmaceutical research, drug discovery, or bioinformatics.
Proven expertise in software engineering best practices, including python package development, DevOps, cloud architectures, CI/CD, and engineering tooling
Hands-on experience handling, processing, integrating, and analyzing large heterogenous data sets data in a drug discovery research environment
Experience with Claude Code or equivalent
Strong publication record with demonstrated contributions to the field
WORK LOCATION: This is a hybrid role requiring you to live within commuting distance and work on-site an average of 2.5 days per week or more as needed.
Relocation support available
Relocation assistance may be available based on business needs and/or eligibility.
Candidates must be authorized to be employed in the U.S. by any employer.
U.S. work visa sponsorship (such as TN, O-1, H-1B, etc.) is not available for this role now or in the future.
Sunshine Act
Pfizer reports payments and other transfers of value to health care providers as required by federal and state transparency laws and implem
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s