Skip to content

Open nowPosted 102 days ago

Lead AI Dataloop and Release Engineer

merlinlabs25 open roles

Pay
$165,000 – $220,000 a year
Where
Remote or Boston or San Francisco Bay Area
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowLead AI Dataloop and Release Engineermerlinlabs · Remote or Boston or San Francisco Bay Area
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on merlinlabs's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.2%30 days
This job: posted 102 days ago

The posting

Merlin (NASDAQ: MRLN) is a publicly traded aerospace and defense company building a non-human pilot to deliver full-stack autonomy for any aircraft from takeoff to touchdown. The Merlin Pilot autonomy system powers a growing range of aircraft and mission profiles and has been proven through hundreds of autonomous flights from Merlin's global flight test facilities, including Kerikeri, New Zealand; Quonset Point, Rhode Island; and soon, Bedford, Massachusetts. Headquartered in Boston, Merlin is expanding its organization to accelerate the development and deployment of its autonomy platform, helping customers solve some of aviation's most pressing challenges, from pilot shortages to improving flight safety. Backed by some of the world's leading investors prior to its public listing, Merlin continues to advance the certification and commercialization of autonomous flight across commercial and defense aviation.

About You:

  • You are a software leader who thrives in enabling deployment of next-gen AI models through a comprehensive data strategy for AI model training, simulation and deployment.
  • You have a technically grounded and thorough appreciation that in autonomous systems, the quality of AI model training, simulation and release infrastructure is inseparable from the performance and safety of what you ship.
  • You've built the data flywheels and worked closely with provisioning training clusters, as well as data-driven, physics-based, and high-fidelity simulators.
  • You know what it takes to engineer the data pipeline that makes simulated environments realistic enough to deploy AI models confidently in safety critical environments like aviation and automotive space.
  • You are organized, methodical, and skilled at building systems that other engineers rely on every day.

Responsibilities:

  • Define and execute a comprehensive data strategy that spans AI model training, simulation, and production deployment across safety-critical autonomous systems.
  • Own the end-to-end data pipeline — from raw collection and labeling through curation, versioning, and delivery — ensuring the reliability and scale that training and simulation workflows demand.
  • Build and maintain data flywheels that continuously improve model performance by closing the loop between deployed system behavior and future training iterations.
  • Collaborate closely with teams provisioning and operating large-scale GPU/TPU training clusters to align data delivery with compute capacity and training schedules.
  • Drive the design and integration of data pipelines that feed data-driven, physics-based, and high-fidelity simulators, ensuring simulated environments are realistic enough to support confident AI model validation.
  • Partner with safety, validation, and certification teams to establish data quality standards and traceability practices that satisfy regulatory requirements in aviation and/or automotive domains.
  • Lead, mentor, and grow a team of data and infrastructure engineers, setting technical direction and fostering a culture of rigor, ownership, and continuous improvement.
  • Define and track KPIs for data pipeline health, simulation fidelity, and model readiness, using these metrics to prioritize investments and communicate progress to senior leadership.

Qualifications:

  • Degree in Computer Science, Artificial Intelligence, Data Science, Computer Engineering, Applied Math, or a related subject.
  • 5+ years of engineering experience, with at least 2 years in a technical leadership role owning data infrastructure, MLOps, or AI platform engineering at scale.
  • Demonstrated experience building and operating data pipelines for AI/ML model training, including dataset management, labeling workflows, and data versioning at production scale.
  • Hands-on experience integrating data systems with large-scale distributed training infrastructure (e.g., GPU/TPU clusters, job orchestration, experiment tracking).
  • Deep understanding of simulation pipelines — including data-driven, physics-based, or sensor-realistic simulators — and how data quality directly impacts simulator fidelity and model transferability.
  • Experience working in or alongside safety-critical domains (autonomous vehicles, aviation, robotics, or similar) with an understanding of what rigor, traceability, and validation mean in that context.
  • Strong systems-thinking mindset: you reason about data quality, pipeline reliability, and infrastructure design as interconnected constraints, not isolated problems.
  • Track record of building platforms and tooling that other engineering teams depend on day-to-day, with a high bar for reliability, documentation, and developer experience.
  • Excellent cross-functional communication skills — able to translate technical data strategy into clear priorities for product, safety, and executive stakeholders.

Nice to Have:

  • Experience with domain randomization, synthetic data generation, or sensor simulation techniques used to bridge the sim-to-real gap in autonomous systems.
  • Familiarity with aviation-specific standards (e.g., DO-178C, DO-254) or automotive safety frameworks (e.g., ISO 26262, SOTIF) as they relate to data and software validation.
  • Prior experience building or scaling data flywheel systems — closed-loop pipelines that feed real-world deployment signals back into training and labeling workflows.
  • Hands-on background with perception, planning, or control model development in autonomous vehicles or UAV/UAS systems.
  • Experience with formal data governance, lineage tracking, or provenance tooling in regulated environments.
  • Contributions to open-source tooling in the MLOps, data engineering, or simulation space.

Logistics:

  • We welcome remote applicants for this role, with a preference for candidates based in or willing to relocate to Boston, MA for a hybrid work schedule at our HQ.

Merlin Labs offers an innovative, entrepreneurial, and team-focused startup environment. We also offer a top-notch benefits package (health, dental, life, flexible paid time off, and 401k with match) and work/life integration. Being part of the Merlin team allows you to become part of a small team that supports professional development while working together to achieve our mission.

Merlin Labs is an equal opportunity employer and values diversity. We do not discriminate on the basis of race, religion, color, national origin, genetic information, sex (including pregnancy), gender, gender identity and expression, sexual orientation, age, marital status, military service or obligation or disability status, or any other characteristic protected by law. All job offers are contingent upon the candidate passing background and reference checks.

At this time, we are unable to provide visa sponsorship or consider candidates who require visa transfers. Applicants must be authorized to work in the United States without the need for visa sponsorship now or in the future.

In compliance with federal law, all persons hired will be required to verify identity and eligibility to work in the United States and to complete the required employment eligibility verification form upon hire.

If you require reasonable accommodation in completing an application, interviewing, completing any pre-employment testing, or otherwise participating in the employee selection process, please direct your inquiries to: [email protected]

Merlin Labs does not accept unsolicited resumes from any source other than directly from candidates.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against merlinlabs's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on merlinlabs's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    merlinlabs's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.