Bilingual Norwegian STEM Expert (PhD) — AI Safety
$77 - $81 / hour
Verified partner opportunity
$60 - $70 / hour
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics.
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.
Responsibilities • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality. • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains. • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking. • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. • Provide structured feedback to improve model alignment and safety performance. • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.
Required Qualifications • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline. • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field. • Excellent written English, critical thinking, and analytical reasoning skills. • Ability to consistently evaluate nuanced and policy-sensitive scenarios.
Preferred Qualifications • Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation. • Familiarity with safety policies, content moderation, or evaluation rubric development. • Experience reviewing complex, high-risk, or ambiguous content.
Why Join? • Shape the safety and behaviour of frontier AI models used by millions worldwide. • Work on challenging, real-world safety evaluations across nuanced and high-impact domains. • Collaborate with leading AI researchers, engineers, and safety teams.
This opportunity may suit professionals with relevant experience in AI evaluation, Biology, Research. Review the official description and requirements before applying.
The listing states $60 - $70 / hour. Confirm the final rate, workload, and payment terms during the official application process.
Use your CV to find relevant roles on HumanitApp. This is separate from a partner application.
Match my CV →Optional role alerts from HumanitApp. Subscribe only if you want weekly emails.
Get weekly role alerts →Before you apply
The listing states $60 - $70 / hour. Confirm the final rate, workload, and payment terms during the official application process.
Yes. The button opens the exact verified Mercorlisting using the referral URL published for this role.
No. HumanitApp independently curates the opportunity. The partner platform manages applications and hiring decisions.
No. Choose the primary Apply action to continue directly without giving HumanitApp your name or email address.