Workplace
On-site
Employment
Internship
Published
Aug 11, 2026
Closes
No date supplied
The role
Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including frontier model training, enterprise adoption, defense applications, and autonomous vehicles. Our mission is to develop reliable AI systems for the world's most important decisions.
Coding is one of Scale's strongest and fastest-moving domains. We built SWE-Bench Pro , a contamination-resistant benchmark of 1,865 long-horizon software engineering tasks across 41 repositories — including a first-of-its-kind private set drawn from proprietary startup codebases — which cut frontier model scores from over 70% on SWE-Bench Verified to roughly 23%, and became a reference point for how the industry measures coding agents. We then extended that foundation with SWE Atlas , an evaluation suite spanning Codebase QnA, Test Writing, and Refactoring, which measures the full engineering loop rather than issue resolution alone. And we are among the largest external contributors to the FrontierBench (Formerly Terminal-Bench ) lineage, contributing more tasks than any other single organization to the launch set.
We're looking for a Senior AI Product Manager to own and scale this Coding portfolio from here. In this role, you will define the strategy, roadmap, and operational excellence of Scale's coding data products, RL environments, and agentic coding evaluations. You will work across AI Product Management, ML Researchers, Engineering, Operations, and Go-To-Market teams to turn our benchmark credibility into a durable, revenue-generating product line that leading labs depend on to train the next generation of software engineering agents.
You will serve as the product owner for coding initiatives, driving task design, environment infrastructure, expert contributor quality, customer adoption, and business impact. You will work directly with leading AI labs
Requirements
Department: Gen AI Product