
MCP & Plug In Connectors Expert | $50-$190/hr Remote
Overview
This role puts your hands-on AI expertise to work. You'll evaluate how well large language models handle real-world, multi-step personal tasks—from planning travel to managing home repairs—by creating prompts, executing workflows, and judging AI outputs. The project relies on advanced users who are comfortable with MCP and plugin/connector tools like Google Drive, Expedia, and Notion. If you're someone who already uses AI heavily in your daily life and can spot where it falls short, this is a chance to shape the next generation of personal AI assistants.
What You'll Do6
- 1Design realistic prompts for complex life tasks across health, travel, dining, home services, and career planning.
- 2Complete those tasks yourself while recording your screen, using your personal plugins and connectors along the way.
- 3Document AI performance: note where it succeeds, fails, or produces unrealistic or unsafe suggestions.
- 4Apply detailed scoring rubrics to judge whether outputs are practical, personalized, and well-reasoned.
- 5Flag missed context, overreach, and incorrect tool usage in model responses.
- 6Collaborate with the team to refine evaluation criteria and improve rubric quality.
Requirements6
- 1Based in the US and able to sign a data-sharing consent form via DocuSign.
- 2Substantial experience with MCP and everyday use of LLM plugins/connectors like Google Drive, Expedia, and Notion (at least a few times per week).
- 3A personal LLM account with roughly six months of active history and heavy usage.
- 4Comfort using AI for multi-step planning, research, and decision-making in your own life.
- 5Strong written judgment—you can explain clearly why an AI output is good, incomplete, unsafe, or unrealistic.
- 6Rubric design and evaluation experience is a big plus, especially 100+ hours on prior rubric projects.
Who Should Apply
You're the kind of person who brings AI into almost every personal decision—trip planning, health research, finding a contractor, or deciding where to eat. You're not just a casual user; you know how to set up connectors and make AI work for your actual life. You can think critically about why a response works or falls apart, and you're excited to help build AI that truly understands personal context, preferences, and tradeoffs.
Salary Insight
Hourly pay ranges from $50 to $190, depending on experience and project tasking.
Location
Required Skills
Application Tip
In your application, share a concrete story of a personal task you planned and executed using an MCP connector or plugin—describe the steps, the tools, and what you learned from the AI's performance.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedOffice-Suite Experts | $70-$100/hr Remote
Mercor is teaming up with a leading AI lab to make frontier models better at generating real-world business documents. We're seeking detail-oriented office-suite experts to evaluate AI-produced files in Excel, Word, and PowerPoint, comparing two candidate outputs and choosing the stronger one. In this fully remote, hourly role, you'll use your own Microsoft Office desktop installation to inspect and grade spreadsheet, document, and presentation artifacts. Your judgment will directly influence how AI handles the polished deliverables that keep companies running.
Peraton
VerifiedSenior AI Integration Developer, MCP & Local LLMs
Senior AI Integration Developer will design and implement an AI assistant for a spectrum monitoring web application at Peraton Labs, serving the Department of Defense. You will build MCP servers and tooling to connect the assistant to application data and workflows, primarily using local models like Ollama in air-gapped environments. You will work with a cross-functional team, influencing AI use cases and ensuring reliable performance. This role requires deep understanding of AI integration, domain context, and full-stack development.
YO AI Labs
VerifiedAI Data Science Expert, Model Evaluation & Prompt Engineering
As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.
Mercor
VerifiedProcess Improvement Expert - AI Evaluator
You evaluate AI-generated documents, spreadsheets, and slide decks against strict quality rubrics. Your feedback drives model improvements for top AI research labs. Collaborate async with a remote team. This role blends deep process improvement expertise with AI training.

Mercor
VerifiedManagement Consulting Expert | $150-$220/hr Remote
This remote role puts your consulting expertise to work behind the scenes of AI development. You’ll partner with a leading AI research organization to evaluate how well AI systems handle real-world consulting tasks—not by producing deliverables, but by defining what top-tier work looks like. You’ll design grading rubrics, score sample outputs, and provide written justifications that help train and calibrate AI models. Expect a high-level, intellectually rigorous project with competitive hourly pay.

Wise Skulls Corp.
VerifiedAI Engineer (LLM Agents & Data Engineering)
Lead design and delivery of AI solutions that scale across multiple platforms. Own the end-to-end pipeline from concept to production while driving innovation in large language models. This role shapes how our systems learn and adapt.