Not Accepting Applications
Mechanistic Interpretability (llms) Machine Learning Expert
This is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability.
Not Accepting Applications
This is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability.
Not Accepting Applications
Contribute published ML research expertise to frontier AI evaluation work.
Not Accepting Applications
Design model evaluations and experiments for a leading AI lab's GenAI team.
Not Accepting Applications
Probe LLM failure modes and edge cases for a leading AI lab's GenAI team.
Not Accepting Applications
Evaluate frontier model outputs and document failure modes with clear written analysis.
Not Accepting Applications
Construct adversarial prompts and rubrics to stress-test large language model safety.
Not Accepting Applications
Apply offensive security expertise to red-team AI systems and build safety datasets.
Not Accepting Applications
Apply cybersecurity software engineering skills to evaluate and harden AI systems against security risks and unsafe outputs.
Not Accepting Applications
Apply advanced biology expertise to evaluate AI outputs and improve model safety on biological reasoning and sensitive knowledge.
Not Accepting Applications
Use advanced chemistry expertise to evaluate AI outputs and strengthen model safety on chemical reasoning and hazardous knowledge.
Not Accepting Applications
Design comprehensive AI training scenarios and test cases to improve model behavior and safety alignment.
Not Accepting Applications
Review AI training scenarios for quality, accuracy, and alignment to ensure high-quality training data standards.
Load more · 22 remaining AI safety expertise is crucial for ensuring artificial intelligence systems behave responsibly and reliably. AI safety training projects rely on human reviewers to identify risks, evaluate unsafe outputs, test edge cases, and guide AI alignment with human values. Without expert oversight, AI systems cannot be deployed safely at scale.
AI safety freelance jobs transform analytical and ethical expertise into high-impact AI training work. Professionals working in AI safety and alignment help improve large language model behavior, reduce bias, and prevent harmful responses. These AI safety projects are remote, well-compensated, and essential to the future of responsible AI development.
Each track explains the work, what it pays, and what qualifications it asks for — then lists every project pooled for it.
Live ai safety projects on PARA AI Labs currently pay between $24/hr and $90/hr, with rates listed upfront on every listing. Specialist experience earns the upper end of the range.
Turing, Handshake AI, Prolific, Mercor, Micro1 and others currently list ai safety projects through PARA AI Labs. Each Apply link goes directly to the hiring platform.
Click Apply on any project below — you'll go straight to the source platform's application, with no middlemen and no fees. Most platforms ask for a short skills assessment before matching you to paid work.
Join thousands of professionals earning from AI training jobs worldwide.