LLM Trainer - Agent Function Call
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
$70–$84/hour
We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics.
Sourced from AIUC via Mercor · original listing · application link last checked 4 Aug 2026
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
Evaluate AI model outputs for accuracy, safety and helpfulness across domains.
This is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability.
Tell us what you know — we'll surface the AI training work that fits.