Senior Software Engineer CMS Data Pipelines
Overview
Design and lead development of Spark-based data pipelines for CMS scoring systems. Build ETL routines and data engineering solutions across Postgres Redshift and S3 Parquet. Collaborate with UI UX and quality teams to define data requirements. Ensure compliance with CMS standards while supporting clinicians. Must reside in the United States and work remotely from any U.S. location.
What You'll Do9
- 1Apply computer science principles to design scalable Spark applications for big data processing
- 2Develop ETL pipelines using Scala APIs and Spark SQL to aggregate government data
- 3Create data models and structures supporting end-to-end data lifecycle management
- 4Implement unit and integration tests for all data processing components
- 5Partner with DevOps engineers to implement CI CD and infrastructure as code practices
- 6Conduct code reviews and establish quality improvement processes
- 7Work with cross-functional teams to translate business requirements into technical specifications
- 8Maintain and enhance Spark ecosystem expertise including Spark Engine and Dataset API
- 9Support CMS scoring initiatives through collaboration with clinical stakeholders
Requirements10
- 1Bachelor’s degree in Computer Science or related field with 5 years progressive software development experience OR Master’s degree with 3 years high-volume Scala Spark experience
- 2Proven SQL development and tuning proficiency with 3 years of experience
- 32 years hands-on experience with AWS services including EMR Redshift CodeBuild Lambda and ECS
- 42 years Git version control and issue tracking familiarity via GitHub and Jira
- 5Prior experience processing Medicare or Medicaid datasets
- 6Federal contracting background preferred
- 7Ability to obtain Public Trust clearance
- 8Must reside in United States for minimum 3 years within last 5 years
- 9Eligibility determined by combined education experience and certifications
- 10Remote work arrangement with mandatory U.S. residency requirement
Salary Insight
$175 - $184k per year
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.
Jefferson University Physicians
VerifiedSenior Data Engineer - AWS & Spark
Lead ownership of large-scale ETL pipelines at Jefferson Health. Build and scale data infrastructure using Python AWS and Spark. Drive performance improvements across clinical and research data platforms. Shape technology roadmap for healthcare analytics. Stand out by delivering measurable impact within 90 days.
Guidehouse Digital, LLC
VerifiedSenior Data Engineer, ETL & Cloud Analytics
You will own the design, development, and optimization of scalable data pipelines and cloud-based analytics solutions for mission-critical projects. This role centers on Python, SQL, and Spark to build modern data ecosystems. You will join a multidisciplinary team of architects, analysts, and engineers, collaborating on cloud platforms like AWS and Azure. Your work will directly support analytics, reporting, and operational workloads. This position requires 3+ years of hands-on experience and offers the chance to lead technical efforts in a federal consulting environment.

MMD Services, Inc
VerifiedSenior Data Engineer MMD Services Cloud Pipeline Specialist
Lead ownership of scalable data infrastructure supporting over one thousand locations. Build robust ETL systems using Spark and Airflow. Drive performance improvements across massive datasets. Shape architecture for future growth. This role impacts enterprise-wide data strategy.

Finoit Inc.
VerifiedSenior Software Engineer Data Infrastructure
Design and lead development of scalable data pipelines for AI/ML platforms. Own design and implementation of distributed systems using Python and cloud services. Drive improvements in data quality and visualization. Lead cross-functional teams to deliver high-performance solutions.
Massachusetts General Physicians Organization, Inc.
VerifiedER - 83026 - CLin - nMD
Lead ETL pipeline initiatives using Python and Spark while collaborating with cross‑functional teams to improve data infrastructure. This role drives data scalability and supports clinical research objectives at Mass General Brigham.

Software Resources, Inc.
VerifiedSenior Data Engineer, Enterprise Data Platform
You will own the design and delivery of scalable data solutions at a major corporation, supporting enterprise data modernization. 5+ years of hands-on data engineering experience with Spark, Airflow, and AWS sets the foundation. You will collaborate with cross-functional teams to build pipelines that power critical business decisions. This hybrid contract role offers a competitive rate of $90-$93/hr.