Senior AI Product Manager, Code

Scale AI Scale AI · Data AI · New York, NY +1 · Gen AI Product

Senior AI Product Manager to own and scale Scale's coding data products, RL environments, and agentic coding evaluations. Define strategy, roadmap, and operational excellence for a revenue-generating product line that leading labs depend on to train software engineering agents. Work with AI Product Management, ML Researchers, Engineering, Operations, and GTM teams.

What you'd actually do

  1. Own the roadmap and strategy for Scale's Coding portfolio, defining priorities across SFT and preference data, reinforcement learning environments, agentic task suites, and evaluation products.
  2. Extend the SWE-Bench Pro and SWE Atlas franchises; deciding what comes next as agents saturate current tasks, and converting benchmark authority into training-data and environment revenue.
  3. Facilitate exploration of the Coding domain and drive alignment among AI-PM, ML, Engineering, Operations, and GTM stakeholders.
  4. Evaluate, prioritize, and operationalize new coding product proposals, ensuring alignment with customer demand, model capability frontiers, and company strategy.
  5. Define and manage the end-to-end coding product lifecycle, from ideation and task taxonomy design to pilot, launch, scaling, and sunset decisions.

Skills

Required

  • product management
  • technical depth in software engineering
  • familiarity with how coding models are trained and evaluated
  • experience building and scaling products that require coordination across engineering, operations, and business teams
  • stakeholder management
  • executive communication skills
  • analytical skills
  • translate ambiguous market and customer signals into clear product strategy

Nice to have

  • AI researchers
  • ML teams
  • developer tools
  • human-in-the-loop data pipelines
  • entrepreneurial mindset
  • bias for action
  • comfort operating in fast-moving, ambiguous environments

What the JD emphasized

  • coding agents
  • agentic coding evaluations
  • agentic task suites
  • coding models are trained and evaluated
  • agentic scaffolds and harnesses
  • coding benchmarks

Other signals

  • coding agents
  • evaluation products
  • training data products
  • RL environments