waypointjobs

Confidential

AI Safety Practitioner for Model Evaluation

Dumfries, VA

Check who can apply and the requirements below before continuing.

Availability awaiting confirmation

We are waiting for a fresh update from the source. This page preserves the last received job details; current availability is not confirmed.

Job description

Help strengthen the safety, quality, and alignment of frontier AI models by evaluating their responses across complex, policy-sensitive, and ambiguous topics. This role focuses on structured assessment and feedback that improves model behavior in high-impact real-world domains.

Key Responsibilities Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.

Review material involving misinformation, political persuasion, self-harm, violence, cyber topics, biosecurity, and other sensitive areas.

Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.

Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.

Provide structured feedback to improve model alignment and safety performance.

Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Qualifications Bachelor''s degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.

At least 5 years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, or a related field.

Excellent written English, critical-thinking, and analytical-reasoning skills.

Ability to consistently evaluate nuanced, policy-sensitive scenarios.

Preferred Qualifications Experience with AI safety, RLHF, SFT, trust and safety, or AI evaluation.

Familiarity with safety policies, content moderation, or evaluation-rubric development.

Experience reviewing complex, high-risk, or ambiguous content.

Work Terms Remote hourly engagement.

Compensation 60 to 70 per hour.

What You Will Contribute Shape the safety and behavior of frontier AI models used by millions of people worldwide.

Work on challenging safety evaluations across nuanced, high-impact domains.

Collaborate with AI researchers, engineers, and safety teams.

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Explore related searches

Current related jobs

SBIOSD

WhatJobs

Neurosurgeon – Endovascular Trained

san diego, CA

compensation: $600,000 - $750,000

Opportunity Overview A busy, established neurosurgical private practice in San Diego, California is seeking a fellowship-trained, endovascular …

Listing review due 2026-10-07View job

CompHealth

WhatJobs

A Locums Dermatologist Is Wanted in Maryland

annapolis, MD

From $225.00 to $300.00 Hourly

When it comes to finding the perfect locums assignment, sometimes it is all about who you know. CompHealth has been around for a long time and ha…

Listing review due 2026-10-07View job

Watson Clinic

WhatJobs

Urologist Opportunity

lakeland, FL

Salary not specified

OWN YOUR PRACTICE- without the start-up costs! Thriving PHYSICIAN OWNED AND OPERATED multi-specialty group is seeking a Urologist to join busy pr…

Listing review due 2026-10-07View job

Allegheny Health Network

WhatJobs

Reproductive Endocrine and Infertility

pittsburgh, PA

Salary not specified

Allegheny Health Network's Department of OB/Gyn is recruiting a fellowship trained REI/IVF physician to join our team in Pittsburgh PA! Job Du…

Listing review due 2026-10-07View job