Retail Specialist, AI Training & Evaluation
Overview
You will guide AI research teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. You will design challenging retail tasks and write accurate solutions grounded in real practice. You will evaluate AI outputs against rubrics and provide feedback. This role offers remote flexibility and pays up to $80/hour. You will collaborate with experts to refine evaluation guidelines and scoring rubrics.
What You'll Do6
- 1Build domain-relevant retail tasks that test AI reasoning in merchandising, category management, and operations.
- 2Write accurate, well-reasoned solutions for retail scenarios based on 10+ years of applied experience.
- 3Evaluate AI model outputs against structured rubrics and score correctness, judgment, and reasoning quality.
- 4Provide clear written feedback on AI answers to improve model training.
- 5Develop and refine evaluation guidelines and scoring rubrics for retail tasks.
- 6Collaborate with subject matter experts to ensure consistency and accuracy in training data.
Requirements6
- 18+ years dedicated retail experience in merchandising, category management, or operations at a top-tier firm like Amazon, Walmart, or Target.
- 2Hands-on experience evaluating LLM/AI outputs against rubrics (mandatory, describe in application).
- 3Demonstrated career progression from Category Manager to Senior Manager to Director level.
- 4Availability for 35+ hours/week during weekdays.
- 5Strong verbal and written communication, problem-solving, and interpersonal skills.
- 6W-2 employment with Cincinnatus LLC.
Salary Insight
$60 - $80k per year
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedRetail Specialist | $60-$80/hr Remote
Work remotely with a top-tier AI lab’s generative AI team to help refine the reasoning and judgment of advanced large language models. As a retail subject-matter expert, you’ll apply your deep industry knowledge to design real-world tasks, evaluate model outputs, and shape the data that powers next-generation AI. This W-2 contract role, facilitated by Cincinnatus LLC, places you directly within a leading AI research environment, offering a rare chance to influence foundational AI development while leveraging your retail expertise.
Mercor
VerifiedRetail SME - AI Evaluation Expert
You will evaluate and refine AI model outputs for retail merchandising, category management, and operations reasoning. Your work drives training data quality for leading AI research labs. Collaborate with engineers and subject matter experts to build rubrics and tasks grounded in real retail practice. This role stands out for its blend of retail expertise and LLM evaluation, offering $60–$80/hour for 8+ years professionals.

Mercor
VerifiedMarketing Specialist | $60-$80/hr Remote
This remote, part-time role places senior marketing experts inside a leading AI lab's GenAI team. You'll help refine how large language models understand and solve real-world marketing challenges. Your daily work involves designing challenging marketing tasks, scoring AI outputs against detailed rubrics, and giving engineers clear feedback to improve model reasoning. Employed by Cincinnatus LLC, you'll be embedded with the AI lab's extended workforce, bringing professional brand and growth marketing judgment into AI training data.

Mercor
VerifiedRetail Banking Expert | $60-$70/hr Remote
Mercor is hiring senior retail banking professionals to help train AI systems that operate in Fortune 500 consumer banking environments. You'll craft realistic enterprise banking scenarios, draft reference outputs, and build scoring rubrics that capture how experienced banking operators make decisions. This is a remote, hourly role paying $60–$70/hr, and it's ideal for anyone with deep hands-on knowledge of large-scale lending, compliance, and branch or digital banking operations. What sets this work apart is the focus on real-world operational judgment rather than textbook theory.

Mercor
VerifiedManagement Consulting Expert | $150-$220/hr Remote
This remote role puts your consulting expertise to work behind the scenes of AI development. You’ll partner with a leading AI research organization to evaluate how well AI systems handle real-world consulting tasks—not by producing deliverables, but by defining what top-tier work looks like. You’ll design grading rubrics, score sample outputs, and provide written justifications that help train and calibrate AI models. Expect a high-level, intellectually rigorous project with competitive hourly pay.

Mercor
VerifiedFinance Specialist | $65-$90/hr Remote
This remote role puts finance expertise at the center of AI development. You'll work with a leading GenAI lab's research team, helping large language models reason through real-world financial scenarios. Your professional experience becomes the yardstick for evaluating model outputs, from valuation logic to risk assessment. The position offers a competitive hourly rate and the chance to shape some of the most advanced AI systems in production.