Skip to content

Open nowPosted 128 days ago

Member of Technical Staff (Post Training)

inherent5 open roles

Where
London
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowMember of Technical Staff (Post Training)inherent · London
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on inherent's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 128 days ago

The posting

MEMBER OF TECHNICAL STAFF, POST-TRAINING — INHERENT (LONDON)

At Inherent, we are on a mission to build AI that recursively self-improves to discover new knowledge. Scientific advances are the backbone of our economic, technological and societal prosperity, but ideas are getting harder to find and breakthroughs are becoming more expensive. We are building a new frontier lab dedicated to developing AI that explores “unknown unknowns” to uncover paradigm-shifting research contributions. Science is a social endeavour, and so our mission is inextricably a human-machine teaming problem. We’re starting by reinventing the AI research factory so that our own agents accelerate their own creation.

Inherent is a well-funded, fast-growing neo-lab backed by Tier 1 VCs who believe in our ethical stance. We are a team of operators with backgrounds at frontier labs who have done foundational work in recursive self-improvement, AI Scientists, world modelling, meta-RL and human-machine cooperation. Working in-person every day at our high-intensity London headquarters, we believe that Europe will lead the way in the coming paradigm of AI-enabled science, unlocking human potential across the globe.

ABOUT THE ROLE

We’re looking for Members of Technical Staff to lead work on post-training state-of-the-art foundation models for open-ended agentic capabilities in scientific research. You’ll be involved at every level of the post-training pipeline: sourcing and creating data, building autocurricula, devising and implementing SFT and RL algorithms, constructing tools and harnesses for foundation model self-improvement, analysing research results, and using information gained to devise future hypotheses. You will work closely with an experienced technical team of humans, and increasingly alongside the AI scientist collaborators we dogfood.

WHAT YOU'D DO

- Design, implement, and tune SFT and RL algorithms to post-train models that autonomously perform state-of-the-art research.

- Build the autocurricula, judges, harnesses and eval pipelines that turn open-ended research tasks into reliable reward signal.

- Run large-scale experiments on state-of-the-art hardware and analyse experiments to determine the next hypotheses to test, in collaboration with our AI agents.

- Close recursive loops so that AI agents drive their own post-training research.

- Work closely with colleagues in the Infrastructure and AI for Science teams to optimise hardware and deliver remarkable performance in real scientific domains.

WHAT WE'RE LOOKING FOR

- 3+ years of deep learning research experience.

- Experience post-training large language, vision, video or multi-modal models.

- Demonstrated track record of success in deep learning research, whether papers, model releases, open-source contributions, or other artifacts.

- 5+ years of software engineering experience, including deep familiarity with Python and at least one deep learning framework (e.g., PyTorch, JAX).

- Experience using the latest coding agents, and opinions about optimal workflow.

- Enthusiasm for experimental organizational design.

- AI-pilled: adopting agents, keen to build a company where agents are front and centre.

STRONG CANDIDATES MAY ALSO HAVE

- PhD in mathematics, computer science or hard science discipline.

- Hands-on experience training LLMs with RL at scale (GRPO/PPO, DPO, distillation, and variants).

- Familiarity with distributed and long-context training infrastructure.

- A background in autocurricula, open-endedness, meta-learning, or recursive self-improvement.

- Experience post-training frontier models at an industry lab (scale, infra, and iteration speed).

WHY THIS IS INTERESTING

- You'll shape the core research of a frontier AI lab from the beginning.

- You'll work on genuine recursive self-improvement — training AI scientists that improve the very pipeline that trains them — not incremental benchmark-chasing.

- You'll dogfood your own work: the agents you post-train accelerate the research that creates them.

- Small team, high trust, no bureaucracy, and a genuinely technical culture.

CULTURE

If you believe in our mission and culture, and are qualified and motivated, we encourage you to apply, even if you don’t meet every one of the criteria above. We know that many of the most creative and talented people have had unusual career paths and backgrounds. Building a team with a diversity of thought is mission-critical, for plurality spurs curiosity, invention and collective experimentation.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against inherent's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on inherent's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    inherent's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.