Skip to content

Open nowPosted 19 days ago

Staff AI Engineer - AI Labs

dlocal51 open roles

Where
Madrid
Work mode
Hybrid
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff AI Engineer - AI Labsdlocal · Madrid
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on dlocal's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.7% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.4%1 day
  2. 3.5%3 days
  3. 7.7%7 days
  4. 13.4%14 days
  5. 34.5%30 days
This job: posted 19 days ago

The posting

Why Join dLocal? dLocal is the financial infrastructure powering global commerce in the world's fastest-growing markets. The biggest companies in the world trust us to unlock growth in 60+ countries across emerging markets—moving money where others see complexity. We don't just process payments; we are architects of payment ecosystems and partners in our customers' expansion. You'll work alongside 1,300+ teammates from 40+ nationalities and tackle global challenges from day one.

What's the Opportunity?

You will join the AI Lab, a team whose mission is to validate high-value emerging AI and automation technologies and de-risk their adoption across dLocal. This is a rare opportunity to work at the frontier of applied AI in fintech: running rigorous experiments on the latest models and tools, and turning results into decisions that shape how a global payments company adopts AI.

This is a senior individual-contributor role. It does not require direct people management, but it carries significant technical influence within the Lab and across the teams that consume its work.

As a Staff AI Engineer in the AI Lab, you own technology scouting, prototyping, and evaluation for dLocal. You will run instrumented spikes and benchmarking on emerging AI technologies, produce clear recommendations for stakeholders across engineering, business, legal and IT based on your prototypes, and coordinate hand-offs to the teams that take validated technologies into production.

For promising technologies, you will help define the patterns, guardrails, and technical requirements needed for adoption and coordinate the hand-off to the engineering teams responsible for productionizing them.

What will I be doing?

  • Run short, instrumented spikes and benchmarking on new models, tools and frameworks: LLMs, agentic systems, vector databases, orchestration frameworks, copilots, assistants and more.
  • Compare vendor and open-source options, documenting trade-offs across quality, cost, latency, security and integration complexity.
  • Deliver concise decision memos with clear recommendations: adopt, watch, or avoid.
  • Design and maintain evaluation environments (e.g. datasets, prompts, scenarios, telemetry) to test models under realistic constraints.
  • Build automation and tooling to measure quality, robustness, latency and cost, including regression tracking over time.
  • Ensure every evaluated technology has benchmark coverage and a documented risk and limitations view.
  • Build enough of a system to understand how a technology behaves under realistic conditions, not just in vendor demos or isolated examples.
  • Explore architecture, integration patterns, operational constraints, security boundaries, and failure modes through working prototypes.
  • Determine what must be true for a proof of concept to become a viable production capability.
  • Prefer focused prototypes that answer specific technical questions over prematurely building production systems.
  • Translate technical findings into clear decision memos for both technical and non-technical stakeholders.
  • For validated technologies, produce readiness guidance covering recommended patterns, guardrails, known limitations, operational considerations, and integration requirements.
  • Coordinate hand-offs to the engineering teams responsible for productionization.
  • Support those teams during the transition when deep context from the evaluation is required, without becoming the permanent owner of the resulting system.
  • Track what happens after Lab recommendations and use those outcomes to improve future evaluation methods.
  • Work with Security, Legal, Compliance and other AI teams to document risk assessments, mitigations and governance recommendations for each evaluated technology.
  • Maintain checklists, decision templates and lightweight standards reusable across evaluations and by partner teams.
  • Incorporate learnings from third-party AI tooling already in use, such as external copilots and the AWS AI suite, into adoption guidelines.
  • Partner with other AI teams and domain teams to ensure clear boundaries and smooth collaboration.
  • Participate in hiring as a technical evaluator and culture champion.
  • Mentor engineers in the Lab and adjacent teams on evaluation methods, benchmarking and experimental design.
  • Share knowledge through internal write-ups, tech talks and occasional external meetups and conferences.

What skills do I need?

  • 8+ years of software engineering experience, including significant experience operating at senior or Staff-level scope.
  • Deep hands-on experience building and evaluating systems based on LLMs and modern AI tooling.
  • Strong software engineering fundamentals and the ability to rapidly build high-quality experimental systems.
  • Experience building agentic or multi-step AI systems involving tool use, orchestration, state, retrieval, or external integrations.
  • Strong knowledge of cloud infrastructure, preferably AWS, and the ability to run experimental workloads securely and cost-consciously.
  • Experience with observability, telemetry, testing, and benchmarking of complex systems.
  • Ability to reason about system architecture, reliability, scalability, asynchronous workflows, and distributed components where relevant.
  • Track record of designing experiments or benchmarks that influenced meaningful technical decisions.
  • Track record designing and running benchmarks that compare AI models and tools under real constraints.
  • Experience constructing evaluation datasets: task selection, labelling, holdout discipline, and keeping a set useful as models improve.
  • Working knowledge of LLM-as-judge methods and their failure modes, alongside human evaluation, inter-annotator agreement, and a view on when each is appropriate.
  • Able to reason about statistical significance on small samples, and to state confidence honestly rather than over-reading a result.
  • Familiarity with regression tracking, telemetry and versioning, so that a result stays reproducible months later.
  • Able to turn ambiguous "we should try this new thing" ideas into well-scoped evaluation plans with clear hypotheses and metrics.
  • Comfortable making trade-off calls across quality, latency, cost and vendor lock-in, and documenting them clearly.
  • Experience writing short, opinionated decision memos that help others move fast.
  • Can explain technical results to non-specialists in concrete, concise terms.
  • Experience working with platform, product and operations teams to align evaluations with real use cases.
  • Able to influence without authority, aligning teams around shared standards and guardrails.
  • Curious and biased toward experimentation, combined with disciplined measurement and risk awareness.
  • Comfortable in a small, high-leverage team without embedded PMs. You structure your own work and keep stakeholders informed.
  • Builder attitude: you prefer reusable tools, templates and playbooks over one-off work.

What do we offer? Besides the tailored benefits we have for each country, dLocal will help you thrive and go that extra mile by offering you: - Flexibility in how you work: We focus on impact and productivity over fixed hours. This means our teams have flexible schedules and, depending on your role and location, you will combine self‑managed focus time with moments of in‑person connection in our collaboration hubs. - Fintech industry: work in a dynamic and ever-evolving environment, with plenty to build and boost your creativity. - Referral bonus program: our internal talents are the best recruiters - refer someone ideal for a role and get rewarded. - Work From Anywhere: Team members can work while traveling for up to 3 months every year.

What happens after you apply?

Our Talent Acquisition team is invested in creating the best candidate experience possible, so don’t worry, you will definitely hear from us. We will review your CV and keep you posted by email at every step of the process!

Also, you can check out our webpage, Linkedin and Youtube for more about dLocal!

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against dlocal's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on dlocal's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    dlocal's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.