Skip to content

Open nowPosted 23 hours ago

Staff+ Research Engineer – Recursive Self-Improvement Lead

Sequen AI16 open roles

Pay
$350,000 – $400,000 a year
Where
United States
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff+ Research Engineer – Recursive Self-Improvement LeadSequen AI · United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Sequen AI's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.3% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.9%1 day
  2. 3.9%3 days
  3. 8.3%7 days
  4. 15.3%14 days
  5. 34.1%30 days
This job: posted 23 hours ago

The posting

STAFF+ RESEARCH ENGINEER — RECURSIVE SELF-IMPROVEMENT LEAD

ABOUT US

Building the ranking intelligence layer for the internet.

Sequen builds Recursive Ranking Intelligence (RRI): an autonomous research engine in which a team of AI agents does the work of an ML research team, autonomously or together with ML researchers, in our clients' own clouds or on Sequen-hosted instances. The models RRI produces serve production traffic for the world's largest retailers, marketplaces and travel platforms, alongside Sequen's ranking platform, which runs frontier ranking models in production at sub-25ms latency and enterprise scale. Each gain compounds into revenue and margin lift measured in hundreds of millions of dollars per customer.

We are a small, highly technical, early-stage team turning recent advances in AI into production systems that operate under unforgiving real-world constraints.

ABOUT THE ROLE

RRI already does ML research on real client problems, and the models it trains serve production traffic at some of the world's largest retailers.

We're looking for a Staff+ Research Engineer to join leadership of our work on recursive self-improvement (RSI): making RRI better at doing research, not only at doing more of it. You will set the research agenda, run the experiments, and ship what works into a system our clients run in production. This is an applied research role: you write the code yourself, and success is measured by what RRI delivers for clients' models. You will join our Recursive Self-Improvement team and help grow it as the area grows. You will work side by side with a Research Engineer focused on the RRI harness, often on the same projects; your profile adds depth in applied research.

KEY RESPONSIBILITIES

- Own the RSI agenda: Define the research roadmap for how RRI improves its own research process, choose the bets, and decide what ships.

- Advance agent strategy: Design and test exploration policies, compute allocation across parallel agents, and agents that revise their own instructions and strategy.

- Learn from past sessions: Turn recorded research sessions into training and evaluation signal.

- Define research quality: Specify what "better research" means and build the eval together with the rest of the RRI team.

- Compound knowledge: Make what RRI learns in one session improve the next, across tasks and clients.

- Drive research frontiers: Track the RSI, autonomous research, and agentic LLM literature, and turn promising ideas into measured experiments within weeks.

- Partner with scientists: Work closely with our applied scientists, who use RRI daily on real client problems, to find where the agents fall short.

ABOUT YOU

- Proven track record: Bring 7+ years of experience in applied ML research or research engineering, with hands-on work in at least one of LLM agents, AutoML, meta-learning, recursive self-improvement, or a closely related area.

- Research into production: Have taken research ideas from prototype to a measured improvement in a shipped product or production system.

- Experimental rigor: Think in baselines, ablations, seeds, and confidence intervals, and distrust any win you have not reproduced.

- Hands-on engineering: Write production-quality Python and build your own experiments end to end.

- Frontier LLM fluency: Understand how frontier models behave in long agentic loops — context limits, compaction, tool use, failure modes, and cost.

- Extreme ownership: Take absolute accountability for a research direction and navigate the ambiguity of an early-stage team.

- Technical leadership: Experience leading or mentoring a small research team.

STRONG CANDIDATES MAY ALSO BRING

- Research footprint: Peer-reviewed publications (e.g., NeurIPS, ICML, ICLR) or widely used open-source work in agents, RL, or AutoML.

- RL and LLM post-training: Experience with reinforcement learning or LLM post-training, such as RLHF, reward modelling, or training agents with RL.

- Evaluation expertise: Experience building evaluations or benchmarks for LLMs or agents, such as reward models, LLM-as-judge, or agentic task suites.

- Ranking domain exposure: Background in search, recommendation, or learning-to-rank models.

- Multi-agent systems: Experience with multi-agent systems or agent swarms, including their cost and safety trade-offs.

WHAT WE VALUE

- Rigorous scientific thinking: You measure before you claim, and you prefer a smaller, reproducible gain to a larger one inside the noise.

- Pragmatic speed: You move from paper to prototype to measured result quickly, without leaving a trail of unrepeatable experiments.

- Research that ships: You judge ideas by what they do for clients' models in production, not by novelty alone.

WHAT WE OFFER

- High-impact influence: A staff-level role that shapes the direction of Sequen's autonomous research engine.

- Pioneering systems: The chance to do frontier RSI research on real production problems, with quantified client outcomes.

- Complete flexibility: Unlimited paid time off, flexible hybrid/remote configurations, and a highly collaborative, world-class engineering culture.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Sequen AI's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Sequen AI's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Sequen AI's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.