Not Accepting Applications
Project Axiom - Axis Sanity Check Rater
Rate and validate axis sanity checks to improve AI model reasoning and evaluation quality.
Not Accepting Applications
Rate and validate axis sanity checks to improve AI model reasoning and evaluation quality.
Not Accepting Applications
Improve AI safety systems by reviewing Spanish-language content.
Not Accepting Applications
Train AI moderation models using Portuguese-language content.
Not Accepting Applications
Review Korean-language content to improve AI moderation and safety systems.
Not Accepting Applications
Train AI moderation systems using Japanese content to detect harmful or policy-violating material.
Not Accepting Applications
Enhance AI safety by reviewing Italian-language content and training moderation systems.
Not Accepting Applications
Improve AI safety models by reviewing Hebrew content and identifying policy violations.
Not Accepting Applications
Train AI systems to detect harmful German-language content and improve moderation accuracy.
Not Accepting Applications
Analyze French content to train AI moderation systems and enhance safety detection.
Not Accepting Applications
Improve AI moderation systems by reviewing Chinese-language content for safety and compliance.
Not Accepting Applications
Train AI moderation systems using Arabic content by identifying harmful patterns and improving safety model accuracy across global platforms.
Not Accepting Applications
Remote red-teaming role focused on probing AI systems for vulnerabilities, misalignment, and safety issues. Help design and execute adversarial prompts and edge cases to improve AI robustness.
Load more · 22 remaining AI safety expertise is crucial for ensuring artificial intelligence systems behave responsibly and reliably. AI safety training projects rely on human reviewers to identify risks, evaluate unsafe outputs, test edge cases, and guide AI alignment with human values. Without expert oversight, AI systems cannot be deployed safely at scale.
AI safety freelance jobs transform analytical and ethical expertise into high-impact AI training work. Professionals working in AI safety and alignment help improve large language model behavior, reduce bias, and prevent harmful responses. These AI safety projects are remote, well-compensated, and essential to the future of responsible AI development.
Each track explains the work, what it pays, and what qualifications it asks for — then lists every project pooled for it.
Live ai safety projects on PARA AI Labs currently pay between $24/hr and $90/hr, with rates listed upfront on every listing. Specialist experience earns the upper end of the range.
Turing, Handshake AI, Prolific, Mercor, Micro1 and others currently list ai safety projects through PARA AI Labs. Each Apply link goes directly to the hiring platform.
Click Apply on any project below — you'll go straight to the source platform's application, with no middlemen and no fees. Most platforms ask for a short skills assessment before matching you to paid work.
Join thousands of professionals earning from AI training jobs worldwide.