Micro1
Micro1Verified Source
RemoteHot

Member of Technical Staff, Research Engineering | $7-$8/hr Remote

7–8/hr
Remote
Posted August 14, 2026
full-time

Overview

A research driven role focused on advancing reinforcement learning systems from concept to production. You’ll build novel RL environments, scalable training pipelines, and automated evaluation tools to push model capabilities. The work blends research ideas with robust, high-performance systems, all remote and full-time.

What You'll Do7

  • 1Create self contained RL environments with reward structures, verifiers, and evaluation logic
  • 2Scale episode pipelines and multi component training processes to support reproducible experiments
  • 3Develop automated data generation using synthetic data to speed up training without sacrificing quality
  • 4Design and integrate AI driven evaluation and QA systems for automated grading, validation, and feedback loops
  • 5Tune and optimize open source RL models using internal datasets and custom training methods
  • 6Set up benchmarking frameworks to measure model ability, robustness, and data quality across tasks
  • 7Contribute to release cycles with evaluations on internal and external benchmarks like micro1 benchmarks

Requirements7

  • 1Strong background in Reinforcement Learning including environment design and training dynamics
  • 2Proven track record in building and scaling RL systems, pipelines, or experimentation frameworks
  • 3Experience with automated data generation and synthetic data pipelines
  • 4Familiarity with automated evaluation, model validation, and quality assurance workflows
  • 5Experience fine tuning and evaluating open source ML models
  • 6Clear, concise communication and solid technical writing skills
  • 7Ability to thrive in fast moving, research oriented, highly collaborative settings

Who Should Apply

Ideal candidates are hands on problem solvers who enjoy bridging research ideas with scalable systems. You should be comfortable working remotely, collaborating across teams, and translating experiments into reliable production components.

Salary Insight

The posting notes a base salary range of $140,000 to $180,000 USD, with equity eligibility and potential performance bonuses.

Location

TypeRemote
LocationRemote
This is a remote position

Required Skills

Reinforcement LearningML-Oriented Data DesignRL EnvironmentsRL Workflows

Application Tip

Prepare a concise portfolio of RL environments or pipelines you built, including metrics that show improvements in training efficiency or model performance.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Micro1

Micro1

3h agoRemotefull-time
Hot

Member of Technical Staff, Frontier AI | $100-$130/hr Remote

A remote senior contributor role focused on bridging research, data, and live AI systems. You’ll own evaluation, failure analysis, and iterative improvements to boost model and system performance. Expect close collaboration with researchers and operators to turn experimental signals into real-world gains and measurable impact. This position emphasizes hands-on ownership, rigorous validation, and clear communication of results across technical and non-technical stakeholders.

100–130/hr
Research Signal JudgmentML-Oriented Data DesignOps-to-Research Translation+1 more
Micro1

Micro1

2h agoRemotefull-time
Hot

Member of Technical Staff, Enterprise AI | $100-$130/hr Remote

A remote, full-time role focused on advancing enterprise AI systems through hands-on research embedded in real workflows. You’ll identify real-world failure modes, run rapid experiments, and translate findings into impactful improvements. Expect to design data and evaluation strategies, build practical tooling, and contribute to external research artifacts that drive robust, scalable AI solutions.

100–130/hr
Research Signal JudgmentML-Oriented Data DesignOps-to-Research Translation+1 more
Bright Vision Technologies

Bright Vision Technologies

14d agoRemotefull-time

Reinforcement Learning Engineer at Bright Vision Technologies

Own reinforcement learning solutions for sequential decision making at scale. Lead design and implementation of RL pipelines for complex environments. Drive adoption of cutting edge RL techniques in production systems. Shape the future of AI-driven products. This role offers rapid growth within a leading consultancy.

100K–150K
PythonDeep Learning FrameworksReinforcement Learning+6 more
Ext

Ext

11d agoRemotefull-time

Principal AI Research Scientist - Alignment - Reinforcement Learning

Lead post-training research at Autodesk AI Lab across London San Francisco Toronto Remote. Own development of model ownership strategies drive model reliability controllability and alignment while shaping architecture decisions across pre training post training and system levels. Impact directly on product deployment and contribute to top conferences.

Competitive salary
Reinforcement LearningRLHFPreference Optimization+2 more

Periodic-Labs

1d agoSan Francisco, Californiafull-time

Research Engineer, AI Model Training

You will own scientific reasoning improvements for frontier models at Periodic-Labs, curating data, building evals, and running large-scale training experiments across thousands of GPUs. You will work with a tight team of physicists, chemists, and reinforcement learning researchers to push the bounds of AI for materials and energy. This role is unique for its focus on midtraining and the foundation it sets for pre-training.

Competitive salary
pythonpytorchaws+2 more
Micro1

Micro1

3h agoRemote
Hot

Senior Software Engineer | $100-$150/hr Remote

A senior software engineer who thrives on building high quality, real world AI training inputs. You’ll shape reinforcement learning environments that test how models handle complex software workflows, from DevOps to debugging, using familiar tools. The role centers on your domain expertise and coding craftsmanship rather than prior AI experience, with a remote contract model.

100–150/hr
· 100 openings
Python3JAVARust+7 more