Skip to content

Open nowPosted 79 days ago

VLA Pre-training Engineer - Deep Learning

humanoid103 open roles

Where
UK, London
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowVLA Pre-training Engineer - Deep Learninghumanoid · UK, London
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on humanoid's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 79 days ago

The posting

Here at Humanoid, we believe in a future where robots amplify human potential. That’s why we’ve set out on a mission to build the world’s most capable, commercially-scalable, and safe humanoid robots. We’re bringing that mission to life with HMND‑01 - our rapidly developed humanoid platform being deployed in real industrial environments - and we’re growing the team to take it even further.

ABOUT THE ROLE

We're hiring a VLA Pre-training Engineer to join our Autonomy team based in London. In this role you will you will work on all aspects of training capable policies, be it pre-training of a base model on a diverse multi-embodiment corpus of trajectories, fine-tuning a policy to perform a specific task well, curating data collection processes or exploring productive ways to generate and use synthetic data. This is primarily a deep learning-focused role, so we are looking for experience solving real problems using modern neural networks, while experience in robotics isn’t strictly required. However if you don’t have such experience, be prepared that you’d need to familiarize yourself with a new domain quickly.

WHAT YOU'LL DO

- Post-train policies via behaviour cloning and RL; own the full loop from data to deployment.

- Partner with the Data Collection team to drive collecting new data: specify what good data looks like, identify failure modes, ensure diversity and coverage.

- Work closely with external partners to ensure steady supply of high-quality pretraining-scale data.

- Run pre-/mid-/post-training on VLA stack; explore new modalities and architecture changes.

- Build and maintain continuous pipelines: ingest synthetic data and teleop logs, version them, apply weak‑supervision labelling, curate balanced datasets, and auto‑surface fresh failure cases into retraining.

- Work with MLOps & Data Platform teams to scale distributed training and optimize models for real‑time edge inference.

WHAT WE'RE LOOKING FOR

- 3+ years building deep‑learning systems (industry or research) with shipped models or published artifacts to show for it.

- Deep hands‑on experience with at least one of: LLMs, VLMs, or image/video generative models — architecture, training, and inference.

- Experience with deep learning infrastructure: streaming datasets, checkpointing & state management, distributed training strategies.

- Strong Python + PyTorch/JAX; you can profile, debug numerics, and write maintainable research code.

- Familiarity with modern software engineering practices.

- You document experiments clearly and communicate trade‑offs crisply.

NICE TO HAVE

- Robotics or autonomous driving experience.

- Experience applying RL to LLMs or robotics.

- Experience with VLA (vision-language-action) models.

- Proven productization of deep nets (latency/throughput constraints, telemetry, on‑device optimization).

- Publications at top-tier deep learning conferences or equivalent open‑source contributions.

- Familiarity with OpenVLA, Physical Intelligence (π) models, or similar open source VLA frameworks.

WHAT WE OFFER

- Competitive equity: stock options with meaningful upside as we scale.

- 30+ paid days off, including 23 days of annual leave, all UK bank holidays, and additional company closure days (including Christmas–New Year shutdown).

- Private healthcare, including virtual and in-person care.

- Pension scheme with 8% total contribution (5% employee, 3% employer) on full earnings.

- Free daily breakfast, catered lunch, and snacks in-office.

- Work at the frontier - collaborate daily with world-class engineers, researchers, and product experts building the next generation of AI and humanoid robotics.

- Real ownership - direct access to founding leadership, meaningful input on product direction, and the ability to drive key initiatives from day one.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against humanoid's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on humanoid's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    humanoid's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.