
Retail Specialist | $60-$80/hr Remote
Overview
Work remotely with a top-tier AI lab’s generative AI team to help refine the reasoning and judgment of advanced large language models. As a retail subject-matter expert, you’ll apply your deep industry knowledge to design real-world tasks, evaluate model outputs, and shape the data that powers next-generation AI. This W-2 contract role, facilitated by Cincinnatus LLC, places you directly within a leading AI research environment, offering a rare chance to influence foundational AI development while leveraging your retail expertise.
What You'll Do5
- 1Partner with research and engineering teams to fill critical gaps in retail merchandising, category management, and operations reasoning.
- 2Create challenging, retail-specific tasks and craft accurate, well-reasoned solutions that reflect authentic industry practices.
- 3Assess AI model responses against structured scoring rubrics, delivering clear written feedback on accuracy, logic, and domain soundness.
- 4Help design and refine evaluation guidelines and scoring criteria tailored to retail scenarios.
- 5Coordinate with fellow subject-matter experts to maintain consistency and high-quality training data across projects.
Requirements5
- 1At least 8 years of professional retail experience in areas like merchandising, category management, buying/planning, or retail operations, ideally at a major recognized company such as Amazon, Walmart, Target, Nike, Costco, or Home Depot.
- 2Proven hands-on experience evaluating AI/LLM outputs against rubrics or structured scoring criteria — this is mandatory and should be clearly detailed in your application.
- 3Visible career growth in retail, for instance progression from Category Manager to Senior Manager or Director of Merchandising.
- 4Dependable availability for at least 35 hours per week during standard weekday hours.
- 5Strong written and verbal communication, sharp problem-solving abilities, and effective collaboration skills.
Who Should Apply
This role is for an experienced retail leader who isn't just deep in the weeds of merchandising or operations but also fascinated by AI and model evaluation. You're comfortable translating your industry judgment into structured feedback and enjoy the precision of working against rubrics. If you've spent years making buying and category decisions at top-tier companies and now want to directly influence how AI understands the retail world, this is your opportunity.
Salary Insight
The role offers an hourly rate of $60.00–$80.00, depending on experience and qualifications.
Location
Required Skills
Application Tip
Make sure to explicitly showcase your experience evaluating AI or LLM outputs in your résumé and cover letter — include specific examples of rubrics you've used, the types of feedback you provided, and how your retail expertise shaped your evaluation. This is the #1 screening criterion.
Similar open positions
Explore active roles that match your skills and interests.
Mercor
VerifiedRetail Specialist, AI Training & Evaluation
You will guide AI research teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. You will design challenging retail tasks and write accurate solutions grounded in real practice. You will evaluate AI outputs against rubrics and provide feedback. This role offers remote flexibility and pays up to $80/hour. You will collaborate with experts to refine evaluation guidelines and scoring rubrics.
Mercor
VerifiedRetail SME - AI Evaluation Expert
You will evaluate and refine AI model outputs for retail merchandising, category management, and operations reasoning. Your work drives training data quality for leading AI research labs. Collaborate with engineers and subject matter experts to build rubrics and tasks grounded in real retail practice. This role stands out for its blend of retail expertise and LLM evaluation, offering $60–$80/hour for 8+ years professionals.

Mercor
VerifiedMarketing Specialist | $60-$80/hr Remote
This remote, part-time role places senior marketing experts inside a leading AI lab's GenAI team. You'll help refine how large language models understand and solve real-world marketing challenges. Your daily work involves designing challenging marketing tasks, scoring AI outputs against detailed rubrics, and giving engineers clear feedback to improve model reasoning. Employed by Cincinnatus LLC, you'll be embedded with the AI lab's extended workforce, bringing professional brand and growth marketing judgment into AI training data.

Mercor
VerifiedRetail Banking Expert | $60-$70/hr Remote
Mercor is hiring senior retail banking professionals to help train AI systems that operate in Fortune 500 consumer banking environments. You'll craft realistic enterprise banking scenarios, draft reference outputs, and build scoring rubrics that capture how experienced banking operators make decisions. This is a remote, hourly role paying $60–$70/hr, and it's ideal for anyone with deep hands-on knowledge of large-scale lending, compliance, and branch or digital banking operations. What sets this work apart is the focus on real-world operational judgment rather than textbook theory.

Mercor
VerifiedLLM Red Team Specialist — Failure Modes & Edge Cases | $60-$90/hr Remote
This role puts you on the front lines of AI safety, working with a top-tier lab to stress-test frontier models and uncover where they fail. You'll design complex, multi-step tasks that probe coding, ML, and reasoning abilities, then transform the weaknesses you find into rigorous evaluation benchmarks. The work is fully remote, runs about 35 hours per week, and pays $60-$90 per hour as a W-2 employee through Cincinnatus LLC. What makes this unique is the tight feedback loop with researchers, giving you direct influence over how the next generation of models is measured and improved.
YO AI Labs
VerifiedAI Data Science Expert, Model Evaluation & Prompt Engineering
As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.