AI Safety Red Teamer
Identify model failure modes through structured AI safety red teaming.
Red-team frontier models: jailbreaks, prompt injection, adversarial testing, and safety evaluation — paid work that makes AI safer.
Identify model failure modes through structured AI safety red teaming.
Lead offensive security red-teaming to evaluate and harden AI systems.
Lead red-teaming quality assurance efforts, identifying vulnerabilities and safety risks in AI model outputs through adversarial evaluation.
Not Accepting Applications
Probe LLM failure modes and edge cases for a leading AI lab's GenAI team.
Not Accepting Applications
Construct adversarial prompts and rubrics to stress-test large language model safety.
Not Accepting Applications
Remote red-teaming role focused on probing AI systems for vulnerabilities, misalignment, and safety issues. Help design and execute adversarial prompts and edge cases to improve AI robustness.
AI red teaming jobs pay you to break models before bad actors do: crafting jailbreaks and prompt injections, probing for bias and misuse, and documenting vulnerabilities so they can be fixed. Alongside them sit AI safety jobs in evaluation — rubric design, scenario writing, and quality control for safety-critical outputs.
Roles run from bilingual red-teaming (adversarial testing in your native language) to security-engineer level adversarial ML — remote, project-based, and in growing demand as every lab ships safety reviews.
One profile covers every partner platform — most people are on a paid ai safety project within a week.
Get matchedAI red teaming jobs pay you to break models before bad actors do: crafting jailbreaks and prompt injections, probing for bias and misuse, and documenting vulnerabilities so they can be fixed. Alongside them sit AI safety jobs in evaluation — rubric design, scenario writing, and quality control for safety-critical outputs.
Live ai safety projects typically pay $48–95/hr, with every rate posted upfront on the listing. Difficulty, seniority, and specialist credentials push rates toward the top of the band.
Yes — every project is remote and asynchronous. Work from anywhere, pick your own hours, no meetings. Apply links go straight to the hiring platform with no middlemen and no fees.
It varies widely: bilingual red-team roles need native language skill and adversarial creativity; advanced tracks want cybersecurity or ML experience. Safety evaluation and scenario design accept strong analytical writers with no security background.