Skip to content

Open nowPosted 186 days ago

Member of Technical Staff – AI Research Engineer (Image/Video Foundation Models)

genpeach4 open roles

Where
Zurich
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowMember of Technical Staff – AI Research Engineer (Image/Video Foundation Models)genpeach · Zurich
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on genpeach's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 186 days ago

The posting

ABOUT GENPEACH AI

GenPeach AI is a product-driven research lab building vertical multimodal foundation models for hyper-realistic human generation in image and video – designed for emotionally resonant, human-centered AI experiences. Our goal is to create tools that supercharge human creativity rather than replace it.

We train models from scratch: proprietary datasets at massive scale, novel architectures and training recipes, large GPU clusters, and tight product integration so research ships to users quickly.

We are a deeply technical team of around 10 people. We’re advised by Directors from Google DeepMind and backed by leading AI-focused funds and angels from OpenAI, Meta AI, Microsoft AI, Project Prometheus, and Fal. Collectively, our team, advisors, and angels have contributed to models including Meta’s Imagine/MovieGen and foundation-model work behind OpenAI’s Sora, plus Google’s Veo and Gemini.

ABOUT THE TEAM

You’ll join the research team working across image/video generation and multimodal understanding. You’ll work closely with other Research Engineers and Scientists, as well as Founders and help turn research into scalable training runs, strong evaluations, and production-ready systems.

ABOUT THE ROLE

We’re hiring an AI Research Engineer to help build and scale GenPeach’s foundation models end-to-end – from implementing new model ideas and training recipes, to owning the parts of the training stack that determine quality and speed, to pushing models through production constraints.

This is a hands-on, high-ownership role. You’ll write research-grade code that becomes production-critical.

IN THIS ROLE, YOU WILL

- Implement and iterate on image/video generative model ideas (architecture, losses, conditioning, sampling, pre-training, distillation, post-training)

- Own training performance end-to-end (distributed training, throughput, memory, stability, debugging scaling failure modes)

- Build the experimentation loop (evals, ablations, reproducibility tooling, reporting, decision hygiene)

- Build and improve VLMs for image/video captioning (data recipes, training strategies, model variants, evaluation)

- Run high-iteration research: read papers when useful, implement ideas, validate empirically

- Create captioning pipelines that improve generation training and product quality

- Partner with inference/product to ship under real constraints (latency, cost, reliability, rollout safety) Build demos and prototypes to showcase capabilities and accelerate iteration

YOU MIGHT THRIVE IN THIS ROLE IF YOU

- Love the craft of experimentation: fast iteration, clear ablations, strong evals, and honest conclusions

- Enjoy debugging messy real-world training runs (not just clean demos)

- Can move between research and engineering: write clean code, ship utilities, and improve team velocity

- Take ownership beyond your job description when needed (startup reality)

- Communicate clearly and collaborate well in a small, senior team

MINIMUM QUALIFICATIONS

- Strong Python and PyTorch skills (4+ years of experience)

- Experience implementing and training deep learning models (generative models, VLMs, LLMs, vision/video, or adjacent)

- Solid understanding of training dynamics, optimization, and practical debugging

- Ability to drive projects end-to-end with minimal supervision

PREFERRED QUALIFICATIONS

- Hands-on experience with diffusion/flow-based image or video generation, or large-scale generative modeling in adjacent domains

- Experience with distributed training at scale (multi-node) and performance tuning (throughput/memory)

- Experience building evaluation frameworks (offline metrics + human eval + regression tracking)

- Strong intuition for data quality and dataset/labeling tradeoffs for training and captioning

- Publications are a plus, but shipped impact and strong technical evidence matter more

WHAT MAKES THIS ROLE UNIQUE

- Build frontier image/video models and the VLM captioning systems that power them

- Join a lean, senior team that holds a high engineering + research bar

- Direct product impact: your training runs become real user-facing capabilities

- Benchmark against the best in the world and compete on model quality through what we ship

HOW WE WORK

- You own outcomes end-to-end and are trusted with real responsibility

- Direct, low-ego communication and fast feedback loops

- Bias toward impact: measure → iterate → ship

- Research discipline: clear ablations, reproducibility, and crisp decision-making

LOGISTICS

- Location: Zurich (Switzerland) or Warsaw (Poland) — onsite or hybrid. If you’re elsewhere, we’re open to remote (team/timezone fit considered).

- Compensation: competitive salary + meaningful equity (level-dependent)

- Interview process: quick screen → 2x technical rounds (practical + systems) → team fit/values

WHAT WE OFFER

- Visa sponsorship (where applicable); we’ll make a strong effort to relocate you to Switzerland or Poland if desired

- Remote-friendly: work fully remote, hybrid, or on-site from our hubs

- Regular offsites and in-person events to collaborate and connect

- Flexible PTO

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against genpeach's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on genpeach's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    genpeach's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.