Scale AI, Inc. · San Francisco, CA · $180600 per hour · Full Time
Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research Organization and will focus on building benchmarks and diagnosing model failure modes in both text and multimodal modalities.
In this role, you will develop rigorous evaluations and diagnostic methods that reveal where frontier models fail and why. You will collabor