Computer Science Expert
Build and review expert computer science tasks used to train and evaluate frontier AI models.
$60–$75/hour
We are building a benchmark dataset to evaluate AI models on professional document understanding and instruction following within the Technology domain.
Tasks consist of complex, multi-step requests grounded in real-world workspace files (technical specs, architecture docs, API references, codebases), web search, and code execution — each paired with a clearly defined ground truth output and an objective evaluation rubric. You will be responsible for authoring tasks that test an AI's ability to reason over technical documentation, follow precise instructions, and produce accurate, well-structured outputs.
We expect a minimum commitment of 15–20 hours per week.
Ideal candidates have 3+ years of hands-on experience in one or more of the following sub-domains:
Sourced from Mercor · original listing · application link last checked 31 Aug 2026
Build and review expert computer science tasks used to train and evaluate frontier AI models.
Build and review expert security engineering tasks used to train and evaluate frontier AI models.
Build and review expert software engineering tasks used to train and evaluate frontier AI models.
Tell us what you know — we'll surface the AI training work that fits.