Skip to content

Open nowPosted 483 days ago

Research Scientist

LAI12 open roles

Where
San Francisco
Work mode
Hybrid
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowResearch ScientistLAI · San Francisco
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on LAI's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.7%30 days
This job: posted 483 days ago

The posting

Research Scientist / Machine Learning Scientist

Location: SF Bay Area/Hybrid / Remote

Type: Full-Time

About the Role:

The Client is seeking a variety of Machine Learning Scientist to help advance how we evaluate and understand AI models. You’ll help design and analyze experiments that uncover what makes models useful, trustworthy and capable through human preference signals. Your work will contribute to the scientific foundations of understanding AI at scale.

This role is deeply interdisciplinary. You’ll work closely with engineers, product teams, marketing and the broader research community to develop new methods for comparing models, analyzing preference data, and disentangling performance factors like style, reasoning, and robustness. Your work will inform both the public leaderboard and the tools we provide to model developers.

If you’re excited by open-ended questions, rigorous evaluation, and research that’s grounded in real-world impact, you’ll find a meaningful home here. We’re looking for:

• Hands-on experience training large-scale models, including reward models, preference models, and fine-tuning LLMs with methods like RLHF, DPO, and contrastive learning.

• Strong foundation in ML and statistics, with a track record of designing novel training objectives, evaluation schemes, or statistical frameworks to improve model reliability and alignment.

• Fluent in the full experimental stack, from dataset design and large-batch training to rigorous evaluation and ablation, with an eye for what scales to production.

• Deeply collaborative mindset, working closely with engineers to productionize research insights and iterating with product teams to align modeling goals with user needs.

Responsibilities:

• Design and conduct experiments to evaluate AI model behavior across reasoning, style, robustness, and user preference dimensions

• Develop new metrics, methodologies, and evaluation protocols that go beyond traditional benchmarks

• Analyze large-scale human voting and interaction data to uncover insights into model performance and user preferences

• Collaborate with engineers to implement and scale research findings into production systems

• Prototype and test research ideas rapidly, balancing rigor with iteration speed

• Author internal reports and external publications that contribute to the broader ML research community

• Partner with model providers to shape evaluation questions and support responsible model testing

• Contribute to the scientific integrity and transparency of the The Client leaderboard and tools

Who is The Client?

Created by researchers from UC Berkeley’s SkyLab https://sky.cs.berkeley.edu/, The Client is an open platform where everyone can easily access, explore and interact with the world’s leading AI models. By comparing them side by side and casting votes for the better response, the community helps shape a public leaderboard, making AI progress more transparent, and grounded in real-world usage.

Why Join Us?

Trusted by organizations like Google, OpenAI, Meta, xAI, and more, The Client is rapidly becoming essential infrastructure for transparent, human-centered AI evaluation at scale. With over one million monthly users and growing developer adoption, our impact is helping guide the next generation of safe, aligned AI systems—grounded in open access and collective feedback.

Our work is regularly referenced by industry leaders pushing the frontier of safe and reliable AI. Sundar Pichai https://x.com/sundarpichai/status/1899779090472644881, Jeff Dean https://x.com/JeffDean/status/1819121162578022849, Elon Musk https://x.com/elonmusk/status/1827055547599712757, and Sam Altman https://x.com/sama/status/1815877987696533897.

• High Impact: Your work will be used daily by the world’s most advanced AI labs.

• Global Reach: Develop data infrastructure powering millions of real-world evaluations, influencing AI reliability across industries at the top-tier

• Exceptional Team: We are a small team of top talent from Google, DeepMind, Discord, Vercel, UC Berkeley, and Stanford.

Requirements:

• PhD or equivalent research experience in Machine Learning, Natural Language Processing, Statistics, or a related field

• Strong understanding of LLMs and modern deep learning architectures (e.g., Transformers, diffusion models, reinforcement learning with human feedback) Proficiency in Python and ML research libraries such as PyTorch, JAX, or TensorFlow

• Demonstrated ability to design and analyze experiments with statistical rigor

• Experience publishing research or working on open-source projects in ML, NLP, or AI evaluation

• Comfortable working with real-world usage data and designing metrics beyond standard benchmarks

• Ability to translate research questions into practical systems and collaborate across engineering and product teams

• Passion for open science, reproducibility, and community-driven research

What we offer:

• The cash compensation for this position has not yet been finalized. Actual compensation will depend on job-related knowledge, skills, experience, and candidate location.

• Competitive salary and meaningful equity

• Comprehensive healthcare coverage (medical, dental, vision)

• The opportunity to work on cutting-edge AI with a small, mission-driven team

• A culture that values transparency, trust, and community impact

Come help build the space where anyone can explore and help shape the future of AI.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against LAI's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on LAI's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    LAI's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.