
Databricks Engineer, AWS & Scala-Spark (Hybrid Boston)
Overview
You will own the design and build of Databricks ingestion and ETL pipelines for a long-term contract in Boston, MA. You will work with AWS, Scala, and Spark in a hybrid model (3-4 days onsite) with an in-person interview required. The role sits within a data engineering team focused on pipeline design and cluster configuration. 7+ years of overall experience in Scala-Spark is expected, and you will have the opportunity to shape data flow architecture from day one.
What You'll Do8
- 1Design and build ingestion pipelines in Databricks that handle batch and streaming data from AWS sources.
- 2Develop ETL jobs using Scala and Spark to transform raw data into analytics-ready datasets.
- 3Configure and manage Databricks workspaces, clusters, and jobs to ensure efficient execution and cost control.
- 4Write and optimize complex SQL queries for data validation and reporting.
- 5Collaborate with data architects to define data models and pipeline standards.
- 6Debug and tune Spark jobs to improve performance and reliability.
- 7Implement cluster autoscaling and job scheduling to meet SLAs.
- 8Document pipeline logic and maintain version control for code and configurations.
Requirements8
- 17+ years of experience in data engineering with a focus on Scala and Spark.
- 2Hands-on Databricks experience including workspace, cluster, and job configuration.
- 3Strong SQL skills for data manipulation and querying.
- 4Proven experience designing and developing ingestion and ETL pipelines.
- 5Experience with AWS services such as S3, EC2, and IAM.
- 6Knowledge of data pipeline orchestration tools like Airflow or similar.
- 7Experience with streaming data processing using Spark Streaming or Delta Live Tables.
- 8Ability to work hybrid onsite in Boston, MA (3-4 days per week).
Salary Insight
$65 - $75k per year
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.

PTR Global
VerifiedDatabricks Data Engineer - New York Contract
Own the design and delivery of scalable data pipelines on Databricks in a contract role based in New York. You will partner with data scientists and analysts to build robust ETL solutions that turn raw data into business-critical insights. This position demands a hands-on engineer who can navigate complex data landscapes with Spark and Delta Lake. Your work will directly influence data-driven decisions across the organization.

Donato Technologies Inc
VerifiedSenior Data Engineer, Databricks & ETL Pipelines
Own the design and delivery of scalable data pipelines on Databricks for a Houston-based enterprise. You join a data engineering team of 6 inside a larger analytics organization, collaborating with architects, analysts, and application teams. Build reliable ETL/ELT solutions that move millions of records daily. This contract role demands hands-on coding in Python and Spark from day one, with direct influence on platform architecture.

Valiant Tek Group, Inc
VerifiedAWS Data Engineer, Databricks & AWS Glue
You will design, build, and optimize enterprise-scale data platforms on AWS for a Santa Monica-based client. Your work directly impacts data pipelines that process petabytes of information across Databricks and AWS Glue. You will join a team of senior engineers and collaborate with data scientists and analysts. This contract role offers the chance to own high-visibility data initiatives from day one.

Protos IT
VerifiedDatabricks Data Engineer, ETL & Lakehouse Architect
Design and own scalable data pipelines on Databricks and Apache Spark for a federal contractor in the DC/MD/VA area. You will modernize legacy ETL into a Delta Lake lakehouse, working cross-functionally with analytics and platform teams. This contract role stands out for its hybrid setup and direct impact on mission-critical data infrastructure.

Nityo Infotech Corporation
VerifiedDatabricks Architect, Data Engineering
You will own the architecture and delivery of Databricks-based data platforms for enterprise clients, defining technical direction across the full project lifecycle. You will design scalable data architectures, lead implementation teams, and drive innovation using Spark, Delta Lake, and cloud-native services. This role sits within a specialized Databricks practice, collaborating with data engineers and solution architects to solve complex data challenges. You will experiment with emerging data technologies and set the standard for best practices in a hands-on, client-facing capacity.
Healthcare Triangle Inc
VerifiedDatabricks Lead - Data Engineering
Own the design and delivery of scalable data engineering solutions on the Databricks platform, driving value across the full data lifecycle. Lead a team of engineers building PySpark, Spark SQL, and Delta Lake pipelines. Collaborate with data architects and stakeholders to shape the lakehouse architecture. This role offers direct ownership of technical strategy and a chance to set standards for data quality and performance.