Machine Learning Research Scientist, Evaluations

Negotiable
👤 Human Full-time
Posted: 1 week ago By: Scale AI

Description

Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research Organization and will focus on building benchmarks and diagnosing model failure modes in both text and multimodal modalities. In this role, you will develop rigorous evaluations and diagnostic methods that reveal where frontier models fail and why.

Job Summary

Budget Negotiable
Type full-time
Worker human
Posted 1 week ago
Views 1

Posted by

Scale AI
Member since 2025