Tech Stack
Solid badges = required, outlined = preferred
Responsibilities
- Conduct AI capabilities research, realistic evaluations, and robust model verification
- Develop high-fidelity, interactive environments capturing the complexity of real-world work
- Evaluate and improve models' ability to plan, execute, and adapt across extended workflows
- Publish research findings, including papers, open datasets, public methodologies, or tools
Soft Skills
Evaluation Methodology
Benefits
- Conference Budget
Culture
In-Person Five Days A WeekFast-Paced
Requirements
Regions: Gb, Us
Get jobs like this in your inbox
Weekly Machine Learning, Artificial Intelligence, Reinforcement Learning hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About mercor
Industry: saas
Size: small
Mercor is a leading AI data company building the layer between human expertise and frontier models, valued at $10 billion.
View company profile →Similar Jobs
Mercor AI Safety Fund Grants
mercor · San Francisco
Mercor Research Fellowship — APEX
mercor · San Francisco
$40k
Research Scientist, APEX Benchmarks
mercor · San Francisco
$15k
Member of Technical Staff, Enterprise Evals Platform
mercor · San Francisco
$15k
Anthropic Fellows Program, AI Safety & Security
Anthropic · London, UK; Ontario, CAN; Remote-Friendly, United States; San Francisco, CA
Remote
$15k