R&D Data Analysis and Machine Learning Software Engineer
Overview
Lead data analysis and machine learning engineer to build and deploy algorithms for environmental research. Own end-to-end pipeline from data ingestion to model deployment. Drive impact across cross-functional teams while collaborating with scientists and engineers.
What You'll Do11
- 1Design and implement scalable data pipelines using Spark and Airflow
- 2Develop machine learning models leveraging TensorFlow and PyTorch
- 3Create technical documentation and present findings at stakeholder meetings
- 4Collaborate with researchers to translate scientific requirements into code
- 5Maintain and optimize production environments using Kubernetes
- 6Review peer code to enhance system architecture and performance
- 7Deploy solutions to cloud infrastructure on AWS
- 8Support ongoing maintenance and iteration of analytical tools
- 9Mentor junior developers and foster knowledge sharing
- 10Ensure compliance with data governance and security standards
- 11Adapt quickly to emerging technologies and research needs
Requirements11
- 1Bachelor’s degree in computer science or related field
- 2Three or more years of experience with regression and machine learning algorithms
- 3Proficiency in Python and MATLAB for data processing tasks
- 4Experience with cloud platforms including AWS services
- 5Strong mathematical foundation with statistical analysis skills
- 6Ability to work independently in high-pressure research environments
- 7Demonstrated leadership in technical project delivery
- 8Commitment to maintaining confidentiality of sensitive information
- 9Excellent written and verbal communication abilities
- 10Willingness to learn new technologies and adapt to evolving research directions
- 11Reliable attendance and punctuality
Salary Insight
Salary not disclosed in listing
Similar open positions
Explore active roles that match your skills and interests.

Finoit Inc.
VerifiedSenior Software Engineer Data Infrastructure
Design and lead development of scalable data pipelines for AI/ML platforms. Own design and implementation of distributed systems using Python and cloud services. Drive improvements in data quality and visualization. Lead cross-functional teams to deliver high-performance solutions.
Voleon
VerifiedSoftware Engineer Strategy Research Analytics
Lead design evolution and long-term architecture of analytics infrastructure supporting research reporting and analysis across strategies. Own critical recurring analytics pipelines and foundational datasets while guiding transition from fragmented bespoke workflows toward a standardized observable query-native platform. Shape technical direction establish reliability standards and drive consolidation efforts improving consistency scalability and reproducibility across research analytics systems.
Matterworks
VerifiedSenior Software Engineer Data Platform
Lead design and scaling of data contracts and infrastructure. Own pipelines serving millions of petabytes. Build scalable label enrichment systems and interfaces for AI and chemistry tools. Ensure operational excellence and data quality at scale.

MMD Services, Inc
VerifiedSenior Data Engineer MMD Services Cloud Pipeline Specialist
Lead ownership of scalable data infrastructure supporting over one thousand locations. Build robust ETL systems using Spark and Airflow. Drive performance improvements across massive datasets. Shape architecture for future growth. This role impacts enterprise-wide data strategy.
4 BMCAP
VerifiedSenior Data Engineer Seattle AWS Spark Airflow Python
Design and lead scalable data pipelines for enterprise analytics. Own development of ETL processes using Python and Spark. Scale infrastructure on AWS. Drive improvements in data quality and performance. This role differs by focusing on cloud-native architecture and cross-functional leadership.

Unisoft Technology Inc
VerifiedSr Lead AI Data Engineer
Lead design and delivery of AI infrastructure to drive scalable machine learning solutions across enterprise platforms. Own development of end-to-end ML pipelines and foster collaboration between data science and engineering teams. This role differs by focusing on cross-functional leadership and production-grade MLOps implementation.