Senior Data Engineer - AWS & Spark
Overview
Lead ownership of large-scale ETL pipelines at Jefferson Health. Build and scale data infrastructure using Python AWS and Spark. Drive performance improvements across clinical and research data platforms. Shape technology roadmap for healthcare analytics. Stand out by delivering measurable impact within 90 days.
What You'll Do9
- 1Design and implement scalable ETL pipelines using Python Spark Airflow
- 2Own data ingestion processes for over 65,000 daily patient interactions
- 3Scale cloud storage solutions leveraging AWS services
- 4Drive migration of legacy systems to modern data architectures
- 5Collaborate with cross-functional teams to define technical requirements
- 6Monitor system performance and optimize resource utilization
- 7Implement data governance frameworks for compliance and quality
- 8Mentor junior engineers in best practices for distributed computing
- 9Develop dashboards for real-time clinical data visualization
Requirements5
- 15+ years building ETL pipelines with Spark and Airflow
- 23-5 seasons leading data engineering initiatives
- 3AWS Certified Solutions Architect Associate
- 4Python programming proficiency
- 5Data orchestration expertise
Salary Insight
Salary not disclosed in listing
Similar open positions
Explore active roles that match your skills and interests.

Inspiration Global
VerifiedSenior Healthcare Data Solutions Architect at Inspiration Global
Design and build enterprise-scale healthcare data solutions driving business outcomes. Lead architecture for complex data systems. Support strategic initiatives in Pittsburgh. Differentiate by delivering scalable data platforms.
4 BMCAP
VerifiedSenior Data Engineer Seattle AWS Spark Airflow Python
Design and lead scalable data pipelines for enterprise analytics. Own development of ETL processes using Python and Spark. Scale infrastructure on AWS. Drive improvements in data quality and performance. This role differs by focusing on cloud-native architecture and cross-functional leadership.
4 BMCAP
VerifiedSenior Data Engineer Seattle AWS Spark Airflow
Design and lead scalable data pipelines for large-scale analytics. Own ETL processes using Python and Spark. Scale infrastructure on AWS. Drive performance improvements. Differentiate by leading cross-functional teams.

RightTalents
VerifiedSenior Data Engineer
Lead design and development of scalable data pipelines. Own end-to-end ETL processes. Drive performance improvements across large datasets. Shape architecture for cloud-native data platforms. Differentiate by delivering production-grade solutions.
Massachusetts General Physicians Organization, Inc.
VerifiedER - 83026 - CLin - nMD
Lead ETL pipeline initiatives using Python and Spark while collaborating with cross‑functional teams to improve data infrastructure. This role drives data scalability and supports clinical research objectives at Mass General Brigham.

MMD Services, Inc
VerifiedSenior Data Engineer MMD Services Cloud Pipeline Specialist
Lead ownership of scalable data infrastructure supporting over one thousand locations. Build robust ETL systems using Spark and Airflow. Drive performance improvements across massive datasets. Shape architecture for future growth. This role impacts enterprise-wide data strategy.