Applied Mathematics Benchmark Specialist
Author and verify rigorous mathematics questions that form gold-standard AI reasoning benchmarks.
Author and verify rigorous mathematics questions that form gold-standard AI reasoning benchmarks.
Write and review business and commerce assessment items used to evaluate AI reasoning quality.
Develop and verify clinical and health science assessment questions to benchmark AI medical reasoning.
Author and review economics and finance assessment items for a frontier AI evaluation benchmark.
Write and verify advanced legal assessment questions used to benchmark AI legal reasoning.
Create and review rigorous biology multiple-choice questions to benchmark AI scientific reasoning.
Author and validate psychology assessment items that establish gold-standard AI evaluation benchmarks.
Write and verify graduate-level history and political science assessment questions for AI benchmarking.
Author and review rigorous philosophy multiple-choice assessment items used to benchmark frontier AI reasoning.
In this hourly, remote contractor role, you will work as an Advertising Subject Matter Expert (SME) to review AI-generated advertising outputs and/or create expert ad copy and campaign assets, evaluating reasoning quality, strategy alignment, and step-by-step creative decision-making while...
If you are an astronomer or space sciences expert who thrives on technical precision, quantitative reasoning, and physics-based problem solving, this is a unique opportunity to contribute directly to how the next generation of AI systems understand and communicate complex space science concepts.
If you are a biomedical engineer who thrives on technical precision, applied problem-solving, and rigorous engineering reasoning in health and life-science contexts, this is a unique opportunity to contribute directly to how the next generation of AI systems understand and communicate complex...
Page 92 of 180 · 2150 total