waypointjobs

Confidential

AI Safety Practitioner for Model Evaluation

Friendship, MD

Check who can apply and the requirements below before continuing.

This job is closed

Applications are no longer available for this announcement. Explore current related opportunities below.

Job description

Help strengthen the safety, quality, and alignment of frontier AI models by evaluating their responses across complex, policy-sensitive, and ambiguous topics. This role focuses on structured assessment and feedback that improves model behavior in high-impact real-world domains.

Key Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.

Review material involving misinformation, political persuasion, self-harm, violence, cyber topics, biosecurity, and other sensitive areas.

Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.

Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.

Provide structured feedback to improve model alignment and safety performance.

Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Qualifications Bachelor''s degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.

At least 5 years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, or a related field.

Excellent written English, critical-thinking, and analytical-reasoning skills.

Ability to consistently evaluate nuanced, policy-sensitive scenarios.

Preferred Qualifications Experience with AI safety, RLHF, SFT, trust and safety, or AI evaluation.

Familiarity with safety policies, content moderation, or evaluation-rubric development.

Experience reviewing complex, high-risk, or ambiguous content.

Work Terms Remote hourly engagement.

Compensation 60 to 70 per hour.

What You Will Contribute Shape the safety and behavior of frontier AI models used by millions of people worldwide.

Work on challenging safety evaluations across nuanced, high-impact domains.

Collaborate with AI researchers, engineers, and safety teams.

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Explore related searches

Current related jobs

Intermountain Health

WhatJobs

CT Technologist $5000 Bonus

murray, UT

up to $3,000 annually

Job Description: Join Our Team as a CT Technologist! We are seeking a dedicated and skilled CT Technologist to join our healthcare team. If …

Last received from source 2026-10-08View job

Intermountain Health

WhatJobs

Mammography Technologist $2500 Bonus

murray, UT

Salary not specified

Job Description: Intermountain Health is seeking a Mammography Technologist to join our Breast Care team at Intermountain Medical Center. This …

Last received from source 2026-10-08View job

FOX Rehabilitation

WhatJobs

Speech Language Pathologist - Port Charlotte, FL

orange county, FL

Salary not specified

Speech Language Pathologist –  Port Charlotte, FL FOX Rehabilitation is growing in Port Charlotte, FL, and we’re looking for passionate, licen…

Last received from source 2026-10-08View job

MedStar Health

WhatJobs

Speech Language Pathologist SLP - Outpatient

sparks, MD

$134,596.00 /Yr

About this Job: MedStar Health is looking for a Speech Language Pathologist (SLP) to join our team at MedStar Franklin Square Medical Center! …

Last received from source 2026-10-08View job