Skip to content

Open nowPosted 24 hours agoWe saw it 85 min after it went up

Research Intern

Mercor114 open roles

Pay
$80 an hour
Where
San Francisco
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowResearch InternMercor · San Francisco
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Mercor's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.0% of postings close within 7 days. Measured by our own scanner across the market. Mercor postings stay open a median of 34 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.5%3 days
  3. 8.0%7 days
  4. 15.0%14 days
  5. 34.1%30 days
This job: posted 24 hours ago

Mercor median: 34 days open

The posting

ABOUT MERCOR

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

ABOUT THE ROLE

As a Research Scientist Intern at Mercor, you’ll work on research at the frontier of post-training, reinforcement learning with verifiable rewards (RLVR), data generation, and model evaluation.

You’ll investigate how datasets, rewards, and training methods affect the capabilities and behavior of large language models. This may include designing controlled experiments, developing new evaluation methodologies, conducting systematic failure analysis, and testing approaches to improve tool use, agentic behavior, and real-world reasoning.

You’ll work closely with research scientists, research engineers, and domain experts to turn open-ended questions into rigorous experiments. Your work will contribute to Mercor’s research agenda and may support external publications, benchmark releases, and the development of frontier AI systems.

WHAT YOU’LL DO

- Develop and investigate research questions related to post-training, RLVR, data quality, and model evaluation.

- Design and run controlled experiments to understand how datasets, rewards, and training strategies affect model performance.

- Study reward-shaping and post-training methods, including approaches such as GRPO and DAPO.

- Develop methods for measuring data quality, usability, and performance uplift on key benchmarks.

- Design and evaluate datasets, rubrics, evaluators, and scoring frameworks for complex model capabilities.

- Conduct systematic error analysis to identify model failure modes and opportunities for improvement.

- Analyze experimental results and communicate findings through clear reports, research artifacts, and presentations.

- Build the research tooling and data pipelines needed to conduct experiments at scale.

- Collaborate with research scientists, research engineers, applied AI teams, and domain experts producing training and evaluation data.

- Contribute to research publications, benchmark releases, and other public research outputs where appropriate.

WHAT WE’RE LOOKING FOR

- Currently pursuing a master’s or PhD in computer science, machine learning, statistics, mathematics, or another relevant field.

- Demonstrated ability to formulate research questions, design experiments, and draw sound conclusions from empirical results.

  • Demonstrated experience in at least one of the following:
  • Training, fine-tuning, or evaluating language models.
  • Agentic AI system, RL environments
  • Developing benchmarks, evaluation methodologies, or data-quality measures.

- At least one publication or open source project.

- Strong programming skills, particularly in Python, and the ability to write reliable research code.

- Familiarity with machine learning fundamentals, experimental design, and statistical analysis.

- Intellectual curiosity.

- Comfort operating in a fast-paced research environment with rapid iteration and a high degree of ownership.

NICE TO HAVE

- Previous research experience in language models, reinforcement learning, model evaluation, or post-training.

- Experience training, fine-tuning, or evaluating language models.

- Familiarity with RLVR techniques, reward modeling, or agentic AI systems.

- Experience developing benchmarks, evaluation methodologies, or data-quality measures.

- Research publications or submissions at competitive CS conferences such as ACL, NeurIPS, ICML, ICLR, or EMNLP.

- Research papers, technical reports, open-source projects, or other work samples demonstrating relevant skills.

Why Mercor

Impact: Your work powers how AI labs train and deploy their models

Learning: Get early exposure to frontier AI research and engineering

Growth: Work with a high-velocity team where interns ship to production

Benefits

- Mentorship from experienced researchers.

- Work on real, high-impact projects.

- $1.5K monthly stipend for meals

- $200 monthly laundry reimbursement

- $200 monthly personal wellness reimbursement

- Free Equinox membership

- Team events and offsites.

- Potential full-time return offer.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Mercor's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Mercor's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Mercor's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.