Mercor
MercorVerified Source
Remote

Retail Specialist | $60-$80/hr Remote

60–80/hr
Remote · United States
Posted August 3, 2026
part-time
10 openings

Overview

Work remotely with a top-tier AI lab’s generative AI team to help refine the reasoning and judgment of advanced large language models. As a retail subject-matter expert, you’ll apply your deep industry knowledge to design real-world tasks, evaluate model outputs, and shape the data that powers next-generation AI. This W-2 contract role, facilitated by Cincinnatus LLC, places you directly within a leading AI research environment, offering a rare chance to influence foundational AI development while leveraging your retail expertise.

What You'll Do5

  • 1Partner with research and engineering teams to fill critical gaps in retail merchandising, category management, and operations reasoning.
  • 2Create challenging, retail-specific tasks and craft accurate, well-reasoned solutions that reflect authentic industry practices.
  • 3Assess AI model responses against structured scoring rubrics, delivering clear written feedback on accuracy, logic, and domain soundness.
  • 4Help design and refine evaluation guidelines and scoring criteria tailored to retail scenarios.
  • 5Coordinate with fellow subject-matter experts to maintain consistency and high-quality training data across projects.

Requirements5

  • 1At least 8 years of professional retail experience in areas like merchandising, category management, buying/planning, or retail operations, ideally at a major recognized company such as Amazon, Walmart, Target, Nike, Costco, or Home Depot.
  • 2Proven hands-on experience evaluating AI/LLM outputs against rubrics or structured scoring criteria — this is mandatory and should be clearly detailed in your application.
  • 3Visible career growth in retail, for instance progression from Category Manager to Senior Manager or Director of Merchandising.
  • 4Dependable availability for at least 35 hours per week during standard weekday hours.
  • 5Strong written and verbal communication, sharp problem-solving abilities, and effective collaboration skills.

Who Should Apply

This role is for an experienced retail leader who isn't just deep in the weeds of merchandising or operations but also fascinated by AI and model evaluation. You're comfortable translating your industry judgment into structured feedback and enjoy the precision of working against rubrics. If you've spent years making buying and category decisions at top-tier companies and now want to directly influence how AI understands the retail world, this is your opportunity.

Salary Insight

The role offers an hourly rate of $60.00–$80.00, depending on experience and qualifications.

Location

Typeremote
LocationUnited States
This is a remote position

Required Skills

retail merchandisingcategory managementretail operationsbuyingplanningllm evaluationrubric developmentai training datalarge language modelsdomain expertisew2remote work

Application Tip

Make sure to explicitly showcase your experience evaluating AI or LLM outputs in your résumé and cover letter — include specific examples of rubrics you've used, the types of feedback you provided, and how your retail expertise shaped your evaluation. This is the #1 screening criterion.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Mercor

15h agoRemotecontract

Retail Specialist, AI Training & Evaluation

You will guide AI research teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. You will design challenging retail tasks and write accurate solutions grounded in real practice. You will evaluate AI outputs against rubrics and provide feedback. This role offers remote flexibility and pays up to $80/hour. You will collaborate with experts to refine evaluation guidelines and scoring rubrics.

60K–80K
Retail merchandisingCategory managementRetail operations+5 more

Mercor

15h agoRemotepayroll

Retail SME - AI Evaluation Expert

You will evaluate and refine AI model outputs for retail merchandising, category management, and operations reasoning. Your work drives training data quality for leading AI research labs. Collaborate with engineers and subject matter experts to build rubrics and tasks grounded in real retail practice. This role stands out for its blend of retail expertise and LLM evaluation, offering $60–$80/hour for 8+ years professionals.

Competitive salary
ai modelllmretail merchandising+2 more
Mercor

Mercor

10d agoRemotepart-time

Marketing Specialist | $60-$80/hr Remote

This remote, part-time role places senior marketing experts inside a leading AI lab's GenAI team. You'll help refine how large language models understand and solve real-world marketing challenges. Your daily work involves designing challenging marketing tasks, scoring AI outputs against detailed rubrics, and giving engineers clear feedback to improve model reasoning. Employed by Cincinnatus LLC, you'll be embedded with the AI lab's extended workforce, bringing professional brand and growth marketing judgment into AI training data.

60–80/hr
· 10 openings
marketingbrand strategygrowth marketing+5 more
Mercor

Mercor

7d agoRemotehourly

Retail Banking Expert | $60-$70/hr Remote

Mercor is hiring senior retail banking professionals to help train AI systems that operate in Fortune 500 consumer banking environments. You'll craft realistic enterprise banking scenarios, draft reference outputs, and build scoring rubrics that capture how experienced banking operators make decisions. This is a remote, hourly role paying $60–$70/hr, and it's ideal for anyone with deep hands-on knowledge of large-scale lending, compliance, and branch or digital banking operations. What sets this work apart is the focus on real-world operational judgment rather than textbook theory.

60–70/hr
· 23 openings
fisfiservtemenos+14 more
Mercor

Mercor

19d agoRemotefull-time

LLM Red Team Specialist — Failure Modes & Edge Cases | $60-$90/hr Remote

This role puts you on the front lines of AI safety, working with a top-tier lab to stress-test frontier models and uncover where they fail. You'll design complex, multi-step tasks that probe coding, ML, and reasoning abilities, then transform the weaknesses you find into rigorous evaluation benchmarks. The work is fully remote, runs about 35 hours per week, and pays $60-$90 per hour as a W-2 employee through Cincinnatus LLC. What makes this unique is the tight feedback loop with researchers, giving you direct influence over how the next generation of models is measured and improved.

60–90/hr
· 10 openings
pythongitmachine learning+9 more

YO AI Labs

14h agoRemotecontract

AI Data Science Expert, Model Evaluation & Prompt Engineering

As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.

Competitive salary
data sciencemachine learningprompt engineering+2 more