WorkaemКарьерная платформа
  • Вакансии
  • Компании
  • Зарплаты
  • Офферы
  • Сервисы
  • Блог
  • Работодателям
Workaem

Карьерная платформа для IT-специалистов: вакансии напрямую с карьерных страниц 300+ компаний, из телеграм-каналов, с международных площадок и от работодателей напрямую. Разбор условий, детектор мёртвых вакансий, AI-инструменты для резюме. Базовые функции бесплатны.

Подпишись, присылаем лучшие вакансии недели
Или читай канал в телеграме
Соискателям
Все вакансииЗа границейУдалёнка в долларахКомпании с РУ основателямиЗарплатыОфферыВозможностиСоветыСоздать резюмеТренировка интервью
По технологиям
Вакансии PythonВакансии JavaScriptВакансии ReactВакансии JavaВакансии GoВакансии Docker
По профессиям
РазработкаДизайнQA / ТестированиеАналитикаProduct / Project ManagerМаркетинг
Работодателям
Разместить вакансиюТарифыБаза кандидатовСвязаться с нами
Кабинет
РегистрацияВойтиЛичный кабинетМои откликиСохранённыеУведомления
Компания
О проектеПредложенияКонтактыБлогКонфиденциальностьУсловия использования
© 2026 Workaem. Все права защищены.КонфиденциальностьУсловияОферта
Made by IT, for IT 💛
Product Manager, APEX
ВердиктОписаниеИнструментыКомпания
  1. Главная
  2. /
  3. Вакансии
  4. /
  5. Product Manager, APEX

Mercor·San Francisco·19 авг.

Product Manager, APEX

🏢 ОфисMiddleПолная занятость
Зарплата не указана
Вилки нет, про деньги придётся договариваться с нуля.
Нажмите на сигнал, чтобы увидеть, на чём он основан

Наша компания

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents. Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

О роли

Mercor’s AI Productivity Index (APEX) assesses how effectively frontier AI models can perform economically valuable work. We're looking for a Product Manager to own and scale the APEX brand and public leaderboard. In this role, you will define the positioning, strategy, roadmap, and operational excellence of Mercor's evaluation products, benchmarks, and public and private leaderboards. You will work across Research, Engineering, Operations, and Go-To-Market teams to transform evaluation datasets into trusted industry benchmarks that influence model development and purchasing decisions across the AI ecosystem You will serve as the product owner for APEX and our eval platform, driving benchmark innovation, evaluation integrity, infrastructure, customer adoption, and business impact. You will work directly with frontier AI labs and enterprise customers, representing Mercor as a thought leader in AI evaluation and measurement. You will partner with Mercor’s world-class benchmark research team, which includes the first authors from many popular benchmarks including Tau Bench, SciCode, PostTrainBench, and more

Чем предстоит заниматься

The ideal candidate combines strong product judgment, technical fluency, operational rigor, and customer-facing experience, with a passion for turning emerging model capabilities into a credible measure of what AI can actually do for the economy
Own the roadmap and portfolio strategy: Set and influence priorities across new benchmark development, leaderboard launches, infrastructure investment, and expansion into new evaluation categories. Decide which domains earn a slot, when a benchmark has saturated, and what replaces it. Identify opportunity areas for strategic partnership
Run the intake for new benchmarks: Evaluate and prioritize proposals from research, customers, and GTM against real demand and company strategy
Enforce eval integrity: Own contamination policy, holdout strategy, versioning, auditability, and release cadence. Publish methodology clearly enough that a skeptical researcher can reconstruct our results
Build end-to-end eval pipeline: Own the path from eval run to published result, including harness execution, grading, model onboarding, hyperparameter scaffolds, cost and latency reporting, and the leaderboard surface itself
Work directly with labs and customers: Understand evaluation needs, present and defend results, and turn what you hear into the next benchmark. Support GTM on launches, partnerships, and thought leadership to influence customer model development strategies
Close the loop to the business: Connect leaderboard demand signals to loss analysis investments and dataset production. Track adoption, usage, and downstream revenue to continuously justify leaderboard ROI
Get in the weeds: Write specs and PRDs, but also read trajectories, spot-check failures, and make small PRs to unblock yourself and continuously improve the system

Наши требования

Experience: 5+ years in product management, technical program management, or a customer-facing technical role. Prior background in SWE, ML, or DS strongly preferred
Eval literacy: You reason fluently about rubric design, inter-rater reliability, agentic harnesses, contamination and overfitting, and what a small sample can and can't support. You can tell a real capability gap from measurement noise
Public judgment: You'll publish numbers about other people's models. You know how to be neutral, precise, and defensible under scrutiny
Taste for what matters: Strong instincts for which capabilities are actually worth measuring, and what results are actually worth highlighting
Entrepreneurial: A track record of building a product, program, or business line from nothing
High Ownership: You take full accountability for outcomes, not just outputs
Independence: Able to self-direct in ambiguous contexts, creating clarity for others
Stakeholder alignment: You can bring research, engineering, operations, and GTM around a shared methodology

Мы предлагаем

Bi-annual performance bonus structure
Generous equity grant vested over 4 years
Up to $15k Relocation bonus
$10K housing bonus (if you live within 0.5 miles of our office)
$1.5K monthly stipend for meals
Free Equinox membership
$200 monthly laundry reimbursement
$200 monthly personal wellness reimbursement
Health, Dental, Vision insurance
M
Mercor
San Francisco

ГрейдMiddle
ЗанятостьПолная занятость
РегионНе Россия
ФорматОфис
ИсточникСкрыто
Опубликовано19 авг.
Все вакансии компании

AI-помощник

под эту вакансию
Войди, чтобы AI оценил твоё соответствие вакансии и написал сопроводительное письмо
Мы против мошенников на площадке: если тебя просят заплатить, продиктовать код или установить непонятное приложение, прекращай общение и сразу пиши нам (чат с основателем или форма обратной связи).