Senior Data Engineer
MCG HealthAbout the role
At MCG, we lead the healthcare community to deliver patient-focused care. We have a mission-driven team of talented physicians and technical experts developing our evidence-based content and innovating our products to accelerate improvements in healthcare. If you are driven to enhance the US healthcare system, MCG is eager to have you join our team. We cultivate a work environment that nurtures personal and professional growth, and this is a thrilling time to become a part of our organization. With dynamic roles that offer meaningful impact, you'll be able to fully realize your potential. Plus, you'll enjoy world-class benefits and the security, stability, and resources of our parent company, Hearst, with over 100 years of experience.
As a Senior Data Engineer you will be responsible for enabling efficient and effective data ingestion & delivery systems. Our team collaborates with data producers (application teams) and data consumers/stakeholders (Data Science, Product, Analytics & Reporting teams) to ensure the availability, quality, and accessibility of data through robust pipelines and storage platforms.
You will:
- Explore, analyze, and onboard data sets from data producers to ensure they are ready for processing and consumption.
- Develop and maintain scalable and efficient data pipelines for data collection, processing (quality checks, de-duplication, etc.), and integration into Data lake and Data warehouse systems.
- Optimize and monitor data pipeline performance to ensure minimal downtime.
- Implement data quality control mechanisms to maintain data set integrity.
- Collaborate with stakeholders for seamless data flow and address issues or needs for improvement.
- Manage the deployment and automation of pipelines and infrastructure using Terraform, Flyte, and Kubernetes.
- Support strategic data analysis and operational tasks as needed.
- Lead end-to-end data pipeline development — from initial data discovery and ingestion to transformation, modeling, and delivery into production-grade data platforms.
- Integrate and manage data from 3+ distinct sources, designing efficient, reusable frameworks for multi-source data processing and harmonization.
What We’re Looking For:
- Demonstrated ability to navigate ambiguous data challenges, ask the right questions, and design effective, scalable solutions.
- Proficient in designing, building, and maintaining large-scale, reliable data pipeline systems.
- Competence in designing and handling large-scale data pipeline systems.
- Advanced SQL skills for querying and processing data.
- Proficiency in Python, with experience in Spark for data processing.
- 3+ years of experience in data engineering, including data modeling and ETL pipelines.
- Familiarity with cloud-based tools and infrastructure management using Terraform and Kubernetes is a plus.
Bonus:
- Experience working with healthcare and clinical data sets
- Experience with orchestration tools like Flyte
Pay Range: $136,000 - $190,400
Other compensation: Bonus Eligible
Perks & Benefits:
💻 Remote work
✈️ Travel expected 2-3 times per year for company-sponsored events
🩺 Medical, dental, vision, life, and disability insurance
📈 401K retirement plan; flexible spending and health savings account
🏝️ 15 days of paid time off + additional front-loaded personal days
🏖️ 14 company-recognized holidays + paid volunteer days
👶 up to 8 w
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s