Machine Learning Research Scientist, Evaluations at Scale AI
- Company: Scale AI
- Location: San Francisco, CA; Seattle, WA; New York, NY
- Employment type: Full-time
- Posted: 2026-08-26
- Technologies: Fine-Tuning
About the role
Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research Organization and will focus on building benchmarks and diagnosing model failure modes in both text and multimodal modalities. In this role, you will develop rigorous evaluations and diagnostic methods that reveal where frontier models fail and why. You will collaborate with researchers and engineers to define best practices in evaluation-driven AI development. You will also partner with top foundation model labs to translate failure analysis into technical and strategic input on the next generation of generative AI models. You will: .…
Apply on Scale AI's official careers page: https://job-boards.greenhouse.io/scaleai/jobs/4728014005