Skip to content

Open nowPosted 79 days ago

Deep Learning Engineer - World Models

humanoid102 open roles

Where
UK, London
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowDeep Learning Engineer - World Modelshumanoid · UK, London
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on humanoid's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market. humanoid postings stay open a median of 1 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.0%30 days
This job: posted 79 days ago

humanoid median: 1 days open

The posting

Here at Humanoid, we believe in a future where robots amplify human potential. That’s why we’ve set out on a mission to build the world’s most capable, commercially-scalable, and safe humanoid robots. We’re bringing that mission to life with HMND‑01 - our rapidly developed humanoid platform being deployed in real industrial environments - and we’re growing the team to take it even further.

OUR MISSION

At Humanoid we strive to create the world's leading, commercially scalable, safe, and advanced humanoid robots that seamlessly integrate into daily life and amplify human capacity.

ABOUT THE ROLE

As a Research Engineer on the World Models team, you will build action-conditioned generative models that predict how the world evolves around our robots — future video, proprioception, contacts, and outcomes — from past observations and actions. World models serve four purposes in our stack: a pretrained, physics-aware prior for our VLA policies; an engine for rare data collection, cross-platform transfer, and sim-to-real transfer; a testbed for policy evaluation and testing before hardware; and a future-prediction rollout engine that surfaces what our policies intend to do, for safety and planning. This is a hands-on individual contributor role: you will design architectures, run large training jobs, and validate your models against real fleet data from industrial deployments.

WHAT YOU'LL DO

- Design and train multimodal world models — video, state, action, and language — using diffusion-based and transformer architectures.

- Build action-conditioned video prediction and dynamics models that stay physically consistent over long horizons, including contact-rich manipulation, and serve as pretrained priors for VLA policies.

- Develop learned-simulator evaluation: score candidate policies offline, predict real-world success rates before deployment, and roll out policy futures to expose intended behaviour for safety review and planning.

- Generate synthetic rollouts and counterfactual experience — including rare events, cross-platform transfer, and sim-to-real transfer — to augment policy training, and measure their effect on downstream task performance.

- Establish fidelity metrics and calibration protocols that quantify where the world model can be trusted and where it diverges from reality.

- Build data pipelines that turn fleet telemetry, teleoperation logs, and internet-scale video into training corpora for world models.

- Run scaling and ablation studies on architecture, data mixture, and context length; communicate findings crisply.

- Collaborate with pretraining, RL, and manipulation teams to integrate world models into policy training and evaluation loops.

WHAT WE'RE LOOKING FOR

- A track record of training large generative models — video, world, or multimodal — with shipped models or published artifacts to show for it.

- Deep hands-on experience with modern generative architectures: diffusion models, autoregressive transformers, latent-variable models, or video prediction.

- Experience with large-scale distributed training: streaming datasets, checkpointing and state management, debugging numerics and training instabilities.

- Strong Python + PyTorch/JAX; you can profile kernels, optimize data loaders, and write maintainable research code.

- Empirical rigor: you design careful evaluations, run honest baselines, and document experiments clearly.

- Excitement about grounding generative models in physical reality rather than pixels alone.

NICE TO HAVE

- Experience with world models for robotics or autonomous driving (e.g., action-conditioned video models, learned simulators, model-based RL).

- Familiarity with robotics simulators (Isaac Sim, MuJoCo) and sim-to-real considerations.

- Experience using world models for policy evaluation or synthetic data generation at scale.

- Publications at top-tier deep learning conferences (NeurIPS, ICML, ICLR, CoRL, CVPR) or equivalent open-source contributions.

- Experience optimizing generative models for fast inference.

WHAT WE OFFER

- Competitive equity: stock options with meaningful upside as we scale.

- 30+ paid days off, including 23 days of annual leave, all UK bank holidays, and additional company closure days (including Christmas–New Year shutdown).

- Private healthcare, including virtual and in-person care.

- Pension scheme with 8% total contribution (5% employee, 3% employer) on full earnings.

- Free daily breakfast, catered lunch, and snacks in-office.

- Work at the frontier - collaborate daily with world-class engineers, researchers, and product experts building the next generation of AI and humanoid robotics.

- Real ownership - direct access to founding leadership, meaningful input on product direction, and the ability to drive key initiatives from day one.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against humanoid's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on humanoid's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    humanoid's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.