
Image Reasoning Generalist | $35/hr Remote
Overview
This remote role puts you at the heart of training advanced AI models by sourcing images that expose their blind spots. You'll focus specifically on counting tasks—finding visuals that trick AI into miscounting, then documenting exactly why. The work is fully remote with a straight $35/hour rate, and it's perfect for detail-obsessed thinkers who enjoy a mix of creative search and structured labeling. No prior AI experience is required, but a sharp eye for visual nuance is essential.
What You'll Do4
- 1Scour the web, your own files, or AI generators for images that could trip up an AI's counting abilities
- 2Pinpoint the specific source of confusion, such as misleading text, repeated objects, lookalike items, or ambiguous measurement units
- 3Test each image against AI models using a standard counting prompt and record the results
- 4Draft extra, more contextual prompts to guide the model, then annotate the task with clear labels
Requirements4
- 1Excellent written English to craft prompts that are unambiguous and easy for models to parse
- 2A meticulous eye for spotting tiny details in images and predicting what might confuse an AI system
- 3Consistency and reliability, since you'll process a high volume of tasks that demand careful judgment
- 4Bonus experience includes photography, image editing, data labeling, or similar detail-focused work
Who Should Apply
You're the kind of person who notices what others miss—whether it's a duplicate object hidden in a photo or an oddly placed caption that throws off interpretation. You enjoy repetitive but mentally engaging work, and you take pride in doing it accurately every time. If you've dabbled in photo editing, data annotation, or simply love puzzles, this role will feel like a natural fit.
Salary Insight
$35 per hour, paid on an hourly basis.
Location
Required Skills
Application Tip
Before applying, create 2–3 sample images that you think would trick an AI counter, and write a short explanation of why. Mention these in your application—it shows you already understand the core challenge.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedAgency Brand Design Expert | $80-$150/hr Remote
This remote opportunity puts your design expertise to work in a new way: instead of creating brand assets, you'll be the standard-setter. You'll define what quality looks like for AI systems by building grading rubrics and scoring real-world design work. Your evaluations will help train and improve AI models used in the creative industry. With a rate of $80-$150 per hour, this freelance role is perfect for senior brand designers who want to influence AI from the evaluator side.
Mercor
VerifiedRetail Specialist, AI Training & Evaluation
You will guide AI research teams to close knowledge gaps in retail merchandising, category management, and operations reasoning. You will design challenging retail tasks and write accurate solutions grounded in real practice. You will evaluate AI outputs against rubrics and provide feedback. This role offers remote flexibility and pays up to $80/hour. You will collaborate with experts to refine evaluation guidelines and scoring rubrics.
YO AI Labs
VerifiedAI Data Science Expert, Model Evaluation & Prompt Engineering
As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.

Mercor
VerifiedManagement Consulting Expert | $150-$220/hr Remote
This remote role puts your consulting expertise to work behind the scenes of AI development. You’ll partner with a leading AI research organization to evaluate how well AI systems handle real-world consulting tasks—not by producing deliverables, but by defining what top-tier work looks like. You’ll design grading rubrics, score sample outputs, and provide written justifications that help train and calibrate AI models. Expect a high-level, intellectually rigorous project with competitive hourly pay.

Mercor
VerifiedAI Safety Practitioner | $60-$70/hr Remote
AIUC is looking for experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models on complex, policy-sensitive topics. In this role, you'll review AI-generated responses for factual accuracy, policy compliance, and overall quality, helping improve how these models handle ambiguous and high-risk grey areas. You'll apply structured evaluation rubrics and provide feedback that directly shapes model behavior. This is a fully remote, hourly freelance position that's ideal for professionals from journalism, policy, or scientific backgrounds who want to apply their expertise to AI alignment.

Mercor
VerifiedGeneralist (Macbook User) | $50-$60/hr Remote
This part-time, fully remote gig places your research and writing skills at the heart of a cutting-edge AI project. You'll help a top AI lab benchmark and refine its latest models by creating questions, answers, and evaluation tools. The work is structured, low-stress, and perfect for postgraduate students or professionals with a few years of experience. You'll need your own Apple Silicon MacBook (M-series with macOS 15+) to participate.