Skip to content

Open nowPosted 8 hours ago

Research Engineer, RL Env

mecka.ai50 open roles

Pay
$170,000 – $300,000 a year
Where
New York
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowResearch Engineer, RL Envmecka.ai · New York
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on mecka.ai's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.1% of postings close within 7 days. Measured by our own scanner across the market. mecka.ai postings stay open a median of 4 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.6%3 days
  3. 8.1%7 days
  4. 15.0%14 days
  5. 33.9%30 days
This job: posted 8 hours ago

mecka.ai median: 4 days open

The posting

ABOUT MECKA AI

Mecka AI is building the data infrastructure layer for robotics and embodied AI. We design and operate global systems for data capture, data labeling, and hardware-enabled workflows used by leading AI labs and robotics companies to train and validate humanoid and embodied AI systems. We work closely with frontier robotics teams to bridge real-world data, simulation, learning-based systems, and deployed hardware.

Build reinforcement learning environments that help researchers train models and understand their capabilities. Across Mecka's Labs team, you'll translate real tasks into computational problems, implement environments for model training and test whether measured progress reflects useful behavior.

This is a hands-on research engineering role for someone who knows how to construct reinforcement learning environments. You'll write working software, investigate failures and develop methods with product colleagues, domain experts and engineers.

WHAT YOU WILL BE DOING:

- Build environments: Define tasks, observations, actions and state transitions. Implement reset behavior, termination conditions and measurable outcomes in environments agents can interact with.

- Develop rewards and scoring: Translate task objectives into feedback and evaluation criteria. Test whether agents can exploit scoring rules without completing the intended task.

- Run learning experiments: Implement baseline agents, train and compare policies, and design controlled experiments that isolate the effects of data, methods and environment changes.

- Make evaluation reliable: Separate training and held-out tasks, check for leakage, version experiments and repeat runs. Report uncertainty and performance across conditions alongside aggregate scores.

- Investigate failures: Inspect trajectories and learning behavior to distinguish policy limitations from data, reward or environment problems. Use findings to prioritize the next experiment.

- Build with the team: Work with domain experts to validate task assumptions and with engineers to turn research prototypes into reusable environments, evaluation tools and documented methods.

WHAT YOU BRING:

- RL environment expertise: You have hands-on experience formulating problems, constructing environments, training agents and critically assessing results.

- Strong programming and software debugging skills; able to build and test research systems that other people can run and extend.

- Sound experimental design and statistical reasoning, including controlled comparisons, evaluation splits, variability and the limits of benchmark results.

- Ability to reason about environment dynamics, reward design and agent behavior, and trace unexpected results to concrete causes.

- Independent research judgment and clear communication; learn unfamiliar domains, work with specialists and explain assumptions, tradeoffs and findings.

EVEN BETTER IF YOU HAVE:

- Experience building interactive environments, simulators or benchmarks used by other researchers.

- Work on agent evaluation, reward design, imitation learning or learning from real-world data.

- Research artifacts with reproducible experiments, useful baselines and evidence of investigating failures beyond headline scores.

A NOTE ON APPLYING

Studies show women and candidates from underrepresented groups often only apply when they meet 100% of the listed qualifications, while others apply after meeting 60%. If you don't check every box above but believe you can do the job, we encourage you to apply — we're looking for capability and trajectory, not a perfect checklist match.

Inclusive Hiring at Mecka

We are committed to creating an inclusive and supportive candidate experience. Should you require any accommodation whatsoever during the interview process, please inform us without any hesitation. Mecka is dedicated to ensuring equal treatment and opportunity in all phases of recruitment, selection, and employment, in compliance with employment law. We do not discriminate based on gender, race, religion, national origin, ethnicity, disability, gender identity/expression, sexual orientation, veteran or military status, or any other protected category. Mecka is proud to be an equal opportunity employer, fostering a culture of inclusivity and maintaining a work environment that is free from discrimination, harassment, and retaliation.

Use of Artificial Intelligence in Recruitment

Mecka uses artificial intelligence (AI) responsibly to support administrative and efficiency-focused aspects of our recruitment process. This includes activities such as drafting job descriptions, generating interview questions, note-taking and recordings, and supporting sourcing and scheduling workflows. All candidate evaluations, interviews, and hiring decisions are made by members of the Mecka team. While AI tools may assist with screening and assessment, they do not replace human judgment in selection decisions. Our use of AI is intended to streamline routine tasks, improve consistency, and enhance the overall candidate experience. We are committed to upholding principles of fairness, transparency, and accountability in all hiring activities. Mecka regularly reviews its recruitment practices to mitigate bias and to ensure alignment with applicable laws and evolving best practices.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against mecka.ai's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on mecka.ai's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    mecka.ai's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.