About this opportunity
Scale lists this GenAI Evaluation Scientist - LLM Benchmarks & Failures opportunity in seattle, Washington. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Scale is seeking a Machine Learning Research Scientist, Evaluations to join the GenAI Research Organization in San Francisco. You will develop rigorous evaluations, diagnose failure modes in frontier LLMs and agents, and design benchmarks for text and multimodal modalities.
Collaboration with researchers and engineers will shape evaluation-driven AI development. The role emphasizes post-training techniques like SFT and RLHF, with opportunities to publish findings at top conferences and influence
#J-18808-Ljbffr
Worksite address
seattle, WA, 98127, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.