Experts

$112/hr

Average contracted rate

444k

Roles created

$4M+

Daily payouts

Research

APEX Benchmarks

The APEX family of benchmarks assesses whether frontier AI models can perform economically valuable tasks across professional services, medicine, and software engineering.

APEX-Agents

Long-horizon, cross-application tasks in professional services

  • Fable 5.1Fable 5.1Max
    68.6%
  • Gemini 3.7 FlashGemini 3.7 FlashHigh
    67.8%
  • Opus 5Opus 5Max
    65.8%
APEX-Accounting

Long-horizon, cross-application tasks in professional accounting

  • Fable 5.1Fable 5.1Max
    61.0%
  • GPT-6 AstraGPT-6 AstraMax
    60.0%
  • Opus 5Opus 5Max
    54.0%
In partnership with Ramp
APEX-SWE

Real-world software engineering across integration and observability

  • Opus 5Opus 5Max
    63.7%
  • Fable 5.1Fable 5.1Max
    63.6%
  • Fable 5Fable 5Max
    58.8%
In partnership with Cognition

Enterprise

Frontier AI, built on your expertise

Mercor helps enterprises turn their expertise into production AI, new revenue streams, and the human intelligence behind the world’s leading models.

Human data & evals

  • 30k+ experts: physicians, lawyers, engineers, consultants.
  • Benchmarks and RL environments built on real professional work (APEX).
  • Human evaluation at the scale frontier labs rely on.

Data monetization

  • Create a new revenue stream from the workflow data your business already generates.
  • Automated anonymization masks 60+ categories of sensitive identifiers.
  • Your organization retains full control over what data is shared.

Enterprise AI

  • Identify high-value AI opportunities by mapping enterprise workflows and institutional knowledge.
  • Deploy production AI with embedded Forward Deployed Engineers and enterprise-grade integrations.
  • Measure and improve AI performance with a platform for context, evaluations, and continuous optimization.