About this opportunity
Scale AI, Inc. lists this GenAI Evaluation Scientist: Benchmarks & Failures opportunity in san francisco, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Scale AI, Inc. seeks Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation.
The role focuses on building benchmarks and diagnosing model failure modes in text and multimodal modalities within the GenAI Research Organization. You will develop rigorous evaluations, collaborate with researchers and engineers, and translate failure analysis into input for next-generation generative AI models.
#J-18808-Ljbffr
Worksite address
san francisco, CA, 94199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.