LLM Trainer - Agent Function Call
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
$17–$25/hour
Location: Remote
Fluent Language Skills Required: English & Indonesian. Native fluency in English and Indonesian is required for this position.
Why This Role Exists
At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who probe AI models with adversarial inputs, surface vulnerabilities, and generate the red team data that makes AI safer for our customers.
This project involves reviewing AI outputs that touch on sensitive topics such as bias, misinformation, or harmful behaviors. All work is text-based, and participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources. Before being exposed to any content, the topics will be clearly communicated.
What You’ll Do
Sourced from Neon via Mercor · original listing · application link last checked 4 Aug 2026
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
Evaluate AI model outputs for accuracy, safety and helpfulness across domains.
This is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability.
Tell us what you know — we'll surface the AI training work that fits.