Skip to content

Open nowPosted 68 days ago

AI Engineer (Search & Matching) - fully remote within Germany (m/f/d) Remote

jobleads15 open roles

Where
Remote DE
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowAI Engineer (Search & Matching) - fully remote within Germany (m/f/d) Remotejobleads · Remote DE
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on jobleads's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 68 days ago

The posting

What it's all about

The Team: JobLeads helps millions of professionals across 40+ countries find their next role, matching them against millions of live jobs. We are building what we believe will be the best job search on the internet - one that understands what a candidate is actually looking for, even when they can't put it into a keyword. Search, matching, and recommendations sit at the core of our product, and applied AI is how we are already winning there.

Your Role: As an Applied AI Engineer, you will build the intelligence behind how candidates and jobs find each other: semantic search, resume-based matching, LLM-driven query understanding, and recommendations.

The role evolves with the mission: whatever it takes to make our search the best job search on the internet is what you'll work on next. That makes out-of-the-box thinking a feature of the job, not a nice-to-have - we expect you to bring new approaches, methods, and frameworks to the table, whether they come from commercial model providers or straight from current research. If a paper published last month suggests a better way to retrieve, rank, or evaluate, we want you to be the person who spots it, tests it, and tells us whether it holds up on our data.

What stays constant is the way of working: form a hypothesis, benchmark it rigorously against the status quo, make it tangible in a prototype people can play with, ship the winner to production, and prove it in an A/B test. You'll work closely with the people who make the decisions and present your results to them directly - your benchmark can become a production experiment within days.

What You'll Be Doing:

  • Design and improve retrieval systems over a large, messy real-world corpus at scale: embeddings, hybrid search, ranking, and re-ranking
  • Shape the design of our vector search backbone: how jobs and resumes are represented, which embedding models to use, the right vector dimensionality for each purpose, etc.
  • Build LLM pipelines for understanding queries, resumes, and jobs — extraction, enrichment, rewriting, matching - choosing the right model and technique for each step
  • Own evaluation as a scientific discipline, not a checkbox:Curate benchmark datasets that capture both the failure modes you're fixing and the healthy cases you must not regress Design LLM-as-judge setups you can defend - ground and calibrate the judge against human sanity checks, detect its biases (a conservative judge will punish a correct query expansion for not being literal), and iterate on the judge itself when it measures the wrong thing Apply and adapt search-relevance methodology - precision@k, NDCG, pairwise preference testing - choosing metrics that resist gaming, and recognizing when a metric rewards degenerate behavior Read results like a scientist: cumulative vs. independent effects, why two valid metrics can disagree, what a distribution's tail says that its mean hides Quantify the full picture - quality, latency percentiles, and cost per query
  • Build interactive prototypes so product managers and stakeholders can test your approaches hands-on before anything ships
  • Partner with engineers to bring winners to production faithfully - no gap between what was benchmarked and what ships - and follow through into A/B test readouts
  • Track the fast-moving model and research landscape and re-evaluate as new models, techniques, and papers appear, keeping us on the best quality-per-euro frontier
  • See your work impact the way millions of job seekers find their next job in 42 countries
  • Be surrounded by colleagues who understand you and help you grow in your role

What You'll Need:

  • Degree in Computer Science, AI, or a related field from a top-tier program (e.g., TUM, University of Tübingen, Saarland University, RWTH Aachen, LMU Munich, TU Berlin, KIT, University of Freiburg, University of Bonn, FAU Erlangen-Nürnberg, TU Darmstadt, University of Stuttgart, Heidelberg University, HPI (Hasso Plattner Institute, Potsdam), TU Dresden, Uni Paderborn, or equivalent)
  • Have hands-on experience building applied ML or LLM systems - retrieval, RAG, agentic pipelines, or traditional search - that real users touched
  • Have genuine data-science depth: you design experiments, build baselines, and reason about metrics and their failure modes before trusting a number
  • Are fluent in Python and comfortable owning a pipeline end to end: data, prompts, embeddings, retrieval, evaluation, and the glue in between
  • Don't trust a demo: you build a baseline, a dataset, and a metric before declaring victory - and you notice when your metric is measuring the wrong thing
  • Communicate clearly with non-ML colleagues and enjoy making your findings understandable, not just correct
  • Are pragmatic and 80/20-minded: you'd rather ship a simple approach that captures most of the gain than a complex one that looks better on paper

Nice If You Have:

  • A track record of taking methods from research papers into working systems - or contributions of your own (publications, open source, competition results)
  • Experience with latency-sensitive production ML and knowing where in a product expensive calls are acceptable
  • Experience running or analyzing A/B tests on search or recommendation systems
  • Exposure to multiple model providers and cost/latency benchmarking across them

What You Can Expect On Board:

  • One of the most exciting applied-AI roles in Germany. This is a rare chance to own the intelligence behind how millions of people find their next job — at real scale, with the latest tooling, and a short path from idea to production.
  • You will never be bored - in the best sense. The problems are at the frontier of applied AI, the loop from idea to live experiment is short, and there's always a harder, more interesting question behind the one you just answered.
  • The best tools, by default. You'll work with the latest AI tooling - Claude Code and whatever comes next - so you move fast, skip the busywork, and spend your energy where your potential actually shows.
  • Friendly, kind, and fast - at a sustainable pace. We move quickly because the work is focused, not because anyone burns out. Communication is direct and warm, and decisions are made on evidence rather than hierarchy: benchmarks settle debates here.
  • Amazing colleagues. Expect good conversations and good moments - experienced product managers, engineers, and data scientists who give real feedback, get genuinely excited about a clean result, and celebrate wins together.
  • Memorable team experiences, from EU meetups to our annual company summer getaway.

We’d love to hear from graduates and PHDs of top-tier AI and Computer Science programs, particularly those with strong machine learning, LLM, or AI systems experience. This includes universities such as TUM, the University of Tübingen, Saarland University, RWTH Aachen, LMU Munich, TU Berlin, KIT, the University of Freiburg, the University of Bonn, and FAU Erlangen-Nürnberg - but outstanding candidates from other universities are equally encouraged to apply.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against jobleads's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on jobleads's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    jobleads's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.