LLM Trainer - Agent Function Call
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
Design and craft multi-turn conversational datasets to improve AI agent reasoning, function calling accuracy, and real-world interaction capabilities.
Evaluate AI model outputs for accuracy, safety and helpfulness across domains.
Evaluate LLM behavior and validate large-scale code repositories.
Evaluate and improve large language models using strong software engineering expertise across multiple programming languages.
Conduct offensive security research to evaluate and harden AI systems.
Identify model failure modes through structured AI safety red teaming.
Assess chemical safety and toxicology risks to strengthen AI model guardrails.
Evaluate personalization quality in AI systems using cultural and contextual understanding of Japanese user behavior.
Evaluate vulnerability-reproduction and remediation tasks used to train frontier AI models.
Join Mercor's SOC investigation specialist network for security AI projects.
Design frontier AI research data and translate operations insight into model improvements.
Advise on child and teen online safety policy to improve AI safeguards.
Page 1 of 5 · 60 total