
Member of Technical Staff, Research Engineering | $7-$8/hr Remote
Overview
A research driven role focused on advancing reinforcement learning systems from concept to production. You’ll build novel RL environments, scalable training pipelines, and automated evaluation tools to push model capabilities. The work blends research ideas with robust, high-performance systems, all remote and full-time.
What You'll Do7
- 1Create self contained RL environments with reward structures, verifiers, and evaluation logic
- 2Scale episode pipelines and multi component training processes to support reproducible experiments
- 3Develop automated data generation using synthetic data to speed up training without sacrificing quality
- 4Design and integrate AI driven evaluation and QA systems for automated grading, validation, and feedback loops
- 5Tune and optimize open source RL models using internal datasets and custom training methods
- 6Set up benchmarking frameworks to measure model ability, robustness, and data quality across tasks
- 7Contribute to release cycles with evaluations on internal and external benchmarks like micro1 benchmarks
Requirements7
- 1Strong background in Reinforcement Learning including environment design and training dynamics
- 2Proven track record in building and scaling RL systems, pipelines, or experimentation frameworks
- 3Experience with automated data generation and synthetic data pipelines
- 4Familiarity with automated evaluation, model validation, and quality assurance workflows
- 5Experience fine tuning and evaluating open source ML models
- 6Clear, concise communication and solid technical writing skills
- 7Ability to thrive in fast moving, research oriented, highly collaborative settings
Who Should Apply
Ideal candidates are hands on problem solvers who enjoy bridging research ideas with scalable systems. You should be comfortable working remotely, collaborating across teams, and translating experiments into reliable production components.
Salary Insight
The posting notes a base salary range of $140,000 to $180,000 USD, with equity eligibility and potential performance bonuses.
Location
Required Skills
Application Tip
Prepare a concise portfolio of RL environments or pipelines you built, including metrics that show improvements in training efficiency or model performance.
Similar open positions
Explore active roles that match your skills and interests.

Micro1
VerifiedMember of Technical Staff, Frontier AI | $100-$130/hr Remote
A remote senior contributor role focused on bridging research, data, and live AI systems. You’ll own evaluation, failure analysis, and iterative improvements to boost model and system performance. Expect close collaboration with researchers and operators to turn experimental signals into real-world gains and measurable impact. This position emphasizes hands-on ownership, rigorous validation, and clear communication of results across technical and non-technical stakeholders.

Micro1
VerifiedMember of Technical Staff, Enterprise AI | $100-$130/hr Remote
A remote, full-time role focused on advancing enterprise AI systems through hands-on research embedded in real workflows. You’ll identify real-world failure modes, run rapid experiments, and translate findings into impactful improvements. Expect to design data and evaluation strategies, build practical tooling, and contribute to external research artifacts that drive robust, scalable AI solutions.
Bright Vision Technologies
VerifiedReinforcement Learning Engineer at Bright Vision Technologies
Own reinforcement learning solutions for sequential decision making at scale. Lead design and implementation of RL pipelines for complex environments. Drive adoption of cutting edge RL techniques in production systems. Shape the future of AI-driven products. This role offers rapid growth within a leading consultancy.
Ext
VerifiedPrincipal AI Research Scientist - Alignment - Reinforcement Learning
Lead post-training research at Autodesk AI Lab across London San Francisco Toronto Remote. Own development of model ownership strategies drive model reliability controllability and alignment while shaping architecture decisions across pre training post training and system levels. Impact directly on product deployment and contribute to top conferences.
Periodic-Labs
VerifiedResearch Engineer, AI Model Training
You will own scientific reasoning improvements for frontier models at Periodic-Labs, curating data, building evals, and running large-scale training experiments across thousands of GPUs. You will work with a tight team of physicists, chemists, and reinforcement learning researchers to push the bounds of AI for materials and energy. This role is unique for its focus on midtraining and the foundation it sets for pre-training.

Micro1
VerifiedSenior Software Engineer | $100-$150/hr Remote
A senior software engineer who thrives on building high quality, real world AI training inputs. You’ll shape reinforcement learning environments that test how models handle complex software workflows, from DevOps to debugging, using familiar tools. The role centers on your domain expertise and coding craftsmanship rather than prior AI experience, with a remote contract model.