Skip to content

Open nowPosted 23 hours ago

Staff Research Engineer – RRI Harness

Sequen AI16 open roles

Pay
$300,000 – $350,000 a year
Where
United States
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Research Engineer – RRI HarnessSequen AI · United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Sequen AI's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.3% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.9%1 day
  2. 3.9%3 days
  3. 8.3%7 days
  4. 15.3%14 days
  5. 34.1%30 days
This job: posted 23 hours ago

The posting

STAFF RESEARCH ENGINEER — RRI HARNESS & APPLIED RESEARCH

ABOUT US

Building the ranking intelligence layer for the internet.

Sequen builds Recursive Ranking Intelligence (RRI): an autonomous research engine in which a team of AI agents does the work of an ML research team, autonomously or together with ML researchers, in our clients' own clouds or on Sequen-hosted instances. The models RRI produces serve production traffic for the world's largest retailers, marketplaces and travel platforms, alongside Sequen's ranking platform, which runs frontier ranking models in production at sub-25ms latency and enterprise scale. Each gain compounds into revenue and margin lift measured in hundreds of millions of dollars per customer.

We are a small, highly technical, early-stage team turning recent advances in AI into production systems that operate under unforgiving real-world constraints.

ABOUT THE ROLE

The harness is the runtime at the heart of RRI. It runs the agent team for hours or days, on GPUs inside our clients' own clouds or on Sequen-hosted instances.

We're looking for a Staff Research Engineer to own the harness and contribute to applied research on how RRI's agents work. This is a research engineering role: you will form hypotheses about agent behaviour, run experiments to test them, and ship what works. You will make RRI reliable, safe and performant across long agent sessions, and extend what it can do. Along the way you will build the eval infrastructure that scores every change to the agents before it ships. You will work side by side with our Recursive Self-Improvement Lead, often on the same projects; your profile adds depth in agent engineering and infrastructure.

KEY RESPONSIBILITIES

- Own the harness: Build and evolve the Python multi-agent runtime — orchestration, tool contracts, context and memory management, compaction, and model adapters across LLM providers.

- Contribute to applied research: Design and run experiments on how the agents work — prompting, tool design, context and memory strategies, and model choice — and ship the changes that measurably improve RRI's results.

- Build eval infrastructure: Deliver a fast, fixed eval set, a scorecard with measured noise bands that runs on every harness change, and component ablations.

- Engineer for resilience: Make multi-day sessions survive GPU failures, node replacement, rolling upgrades, provider errors, and credit limits without a human restart.

- Deliver observability: Build event streams, cross-session analysis, per-agent cost and token accounting, and admin tooling over MCP that let us see inside every session.

- Cut waste: Find where agents idle, loop, or burn tokens, and fix it in code.

- Design clean contracts: Work with the owners of our Go control plane and GPU proxy on the interfaces between components.

- Support clients: Partner with applied scientists and forward-deployed engineers to diagnose and fix issues clients hit in production.

ABOUT YOU

- Proven track record: Bring 7+ years of experience building production software, with strong Python and a history of distributed or long-running systems.

- Agentic LLM experience: Have built with LLMs in agentic loops — tool calling, prompt and context management, streaming, retries — and know how they fail at scale.

- System-wide comfort: Debug confidently across process, container, and network boundaries, on Kubernetes and GPU nodes.

- ML fluency: Read a training script, understand a ranking metric, and tell a real regression from noise.

- Extreme ownership: Take absolute accountability for a system end to end, with a deep care for reliability and observability.

STRONG CANDIDATES MAY ALSO BRING

- Applied research experience: Have run ML or LLM experiments end to end, with baselines and ablations, and turned the results into product changes.

- Evaluation infrastructure: Have built test harnesses, benchmarks, experiment tracking, or CI for ML systems.

- Polyglot capabilities: Working knowledge of Go for contributions to the control plane and proxy.

- Customer-controlled deployments: Experience shipping software into on-prem, BYOC, or air-gapped environments and their security constraints.

- ML platform depth: Experience with PyTorch training at scale, GPU scheduling, or ML platforms.

- Open-source footprint: Contributions to agent frameworks, eval harnesses, or ML tooling.

WHAT WE VALUE

- Rigorous systems thinking: You design for the failure case first, because in a multi-day session every rare failure eventually happens.

- Measure, then change: You would rather ship a smaller change with an eval score than a larger one on intuition.

- Pragmatic speed: You move quickly without accumulating debilitating technical debt.

WHAT WE OFFER

- High-impact influence: A staff-level role that owns the core runtime of Sequen's autonomous research engine.

- Pioneering systems: The chance to build production infrastructure for long-horizon AI agents that train models serving enterprise traffic.

- Complete flexibility: Unlimited paid time off, flexible hybrid/remote configurations, and a highly collaborative, world-class engineering culture.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Sequen AI's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Sequen AI's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Sequen AI's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.