Human intelligence for AI training
Mercor says it powers frontier research, RLHF data, and AI agent training at scale, using human expertise for tasks that require domain judgment.
Mercor is an AI platform for frontier research, human-in-the-loop training data, enterprise agents, and benchmark-driven evaluation for AI labs and enterprises.
Mercor is a platform focused on organizing human intelligence for AI work. Across its site, it positions itself around frontier research, RLHF data, AI agent training, and enterprise AI systems that are grounded in human expertise.
The company’s product surface spans three main areas: research and evaluation for frontier models, enterprise agent deployment for business workflows, and benchmarks such as APEX that measure model and agent performance on economically valuable tasks. The site also shows active hiring and a large contractor network supporting those services.
Mercor says it powers frontier research, RLHF data, and AI agent training at scale, using human expertise for tasks that require domain judgment.
Mercor Enterprise begins by capturing company workflows and tribal knowledge, then uses that input to identify high-value agents and a roadmap for deployment.
The Enterprise offering builds agents around a company’s own definition of quality work, then deploys them to work, learn, and continuously improve.
Mercor Enterprise includes benchmarking to pressure-test AI products with independent, repeatable evaluations and clear evidence of success and failure modes.
Mercor Research produces frontier data, RL environments, benchmarks, peer-reviewed research, evaluation leaderboards, datasets, and open tooling.
APEX benchmarks measure AI performance on economically valuable tasks across professional services, medicine, software engineering, and consumer activities.
Teams can ingest ATS data and role rubrics to screen large applicant pools, then return calibrated scorecards and pass/fail recommendations while keeping recruiters in the loop.
Engineering or operations teams can use the platform to detect incidents, generate runbooks, identify root causes, and reduce time to mitigation across observability and incident tools.
Consulting, finance, and strategy teams can synthesize information from internal and external research sources into investment theses, competitive research, or market-entry materials.
Sales, legal, compliance, and IT teams can combine inputs to produce first-draft RFP responses without manually chasing context across departments.
Organizations can benchmark their AI products or agents with independent evaluations that show where systems succeed, fail, or need improvement.
Mercor appears to focus on human-in-the-loop AI work: frontier research, RLHF data, AI agent training, and enterprise AI agents built around company-specific standards.
The site positions Mercor for top AI labs and enterprises, and also highlights use by research teams, engineering teams, recruiting, consulting, finance, sales, and business operations teams.
The Enterprise page describes a workflow that can include diagnosing workflows, deploying agents, benchmarking AI products, and monetizing enterprise data. The Research page focuses on frontier data, RL environments, benchmarks, and evaluation tooling.
The pricing page is not available and returns a 404, so the site does not expose pricing details in the provided sources.