Skip to content

Open nowPosted 4 days agoWe saw it 8 min after it went up

AI Automation Engineer, Real-World Test Lab

Niantic17 open roles

Pay
$158,400 – $210,000 a year
Where
San Francisco, CA
Work mode
Hybrid
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowAI Automation Engineer, Real-World Test LabNiantic · San Francisco, CA
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Niantic's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market. Niantic postings stay open a median of 27 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 4 days ago

Niantic median: 27 days open

The posting

About Niantic Spatial

At Niantic Spatial, we're building the future of physical AI. Powered by a proprietary database of over 30 billion posed images, our groundbreaking mapping technology unlocks a new dimension of interaction and spatial intelligence that helps both humans and machines better understand, represent, navigate, and engage with the real environment.

Our reconstruction technology captures environments with geometric accuracy and extreme detail from any standard camera, and our Visual Positioning System delivers precise positioning almost anywhere in the world. We serve customers across robotics, the public sector, and energy and industrial markets - building for the 80% of economic activity that takes place beyond our screens.

About the Real-World Test Lab

Physical AI doesn't get graded on a leaderboard. It gets graded on a factory floor at shift change, in a substation with no GPS, on a site where the lighting is wrong and the stakes are real.

The Real-World Test Lab closes the gap between the benchmark and the field. We bring the customer's world inside our walls - their devices, their environments, their hardest conditions, and their definition of success - and make it the bar every release has to clear.

As Niantic Spatial's first and most demanding customer, we push our reconstruction, localization, and spatial understanding to their limits, find where they shine and where they break, and turn that into evidence that shapes what we build next. It's a new team at the frontier of physical AI, and you'll help invent how the job is done.

About the Role

We're hiring an AI Automation Engineer, reporting to the Director of the Real-World Test Lab, to build the system that produces our evidence. Today our evaluations are manual, inconsistent, and slow. You'll design and own the automation that runs them end to end on every relevant release, with no human driving it, turning one-off experiments into an always-on service the whole company relies on.

This role is about owning the evidence, not executing a test plan. You'll decide how each workflow gets exercised, and you'll be measured on whether the company can trust the results, not on how much automation exists. The work counts when it keeps running correctly months later, without you in the loop.

You believe evaluation is engineering, not process. You've built systems that test other systems, and you know the difference between a script that works on your laptop and infrastructure a team can trust. When a result couldn't be reproduced, you fixed the tooling instead of arguing about the number.

What You'll Do

- Build the Evaluation Machine - Own the automation that executes Lab evaluations end to end: environment setup, run orchestration, artifact capture, and result collection. Make reruns free so we test constantly rather than occasionally.

- Make Results Comparable - Instrument scorecards so a result can be compared across product versions, devices, and capture conditions. A number without its lineage is not evidence.

- Automate the Agent-Driven Layer - Build the agent workflows that exercise priority customer use cases at realistic scale, across the graded difficulty spectrum from easy to frontier, and be honest about where agents cannot yet replace human judgment.

- Kill Manual Work Permanently - Convert one-off experiments into standing protocols that run on every relevant release. Automate recurring inspection wherever it can be automated, and route what genuinely cannot to scalable human review.

- Make Failures Actionable - Produce diagnostics precise enough that a finding reaches its owner with a reproducible case and data attached. Findings that need re-investigation before anyone can act on them are half-finished.

- Keep Data From Being the Bottleneck - Work with the AI Data Manager so every run is reproducible from a known dataset state, without anyone downloading and re-uploading data by hand.

What Success Looks Like

- 30 days: One priority workflow evaluated end to end with no manual steps, with results in a comparable scorecard.

- 60 days: Evaluations trigger automatically on relevant releases, with diagnostics that route failures to the right owners.

- 90 days: Three priority workflows under standing automated evaluation, with version-over-version comparison for leadership.

What You'll Bring

- Built and maintained production-grade automation or test infrastructure that other engineers relied on daily.

- Experience evaluating systems where correctness is graded rather than binary, meaning quality, accuracy, or latency thresholds rather than pass/fail assertions.

- Strong Python, with fluency in orchestration, CI-style pipelines, cloud storage, and reproducible environments.

- Worked with backend services and APIs you did not own, integrating against them without becoming a bottleneck for their team.

- Written up a technical finding clearly enough that a non-author could act on it without a meeting.

- A bachelor's degree in a relevant field, or equivalent experience.

Nice to Have

- Built agent-based or LLM-driven automation for a task that previously required human judgment.

- Worked on evaluation or benchmarking for computer vision, 3D reconstruction, or spatial systems.

- Instrumented dashboards or scorecards that leadership used for release decisions.

- Operated data provisioning or registry infrastructure across multiple environments and access models.

Competencies

- Intellectually honest. You design tests that produce answers people can trust, and you refuse to let a benchmark imply more than the evidence supports — including when the honest answer is inconvenient for your own work.

- Pragmatic. You find the fastest path to a defensible result without going through one-way doors that undermine scaling. You have no patience for building infrastructure nobody has asked for yet.

- AI-forward by instinct. You reach for agents and LLM-driven automation before you reach for headcount, and you're rigorous about verifying what they produce.

- Builds for other people. You've maintained something after the initial excitement wore off. Your tooling is documented, your failures are legible, and colleagues use what you build without asking you to run it for them.

Compensation & Benefits

Base salary range of $158,400 to $210,000 per year. Compensation also includes an annual bonus, equity, and a comprehensive benefits package including medical, dental, and vision coverage, 401(k), and more.

Location & Work Model

This role is based in our San Francisco office, with three days per week in office.

Inclusive Application

We know the strongest candidates don't always tick every box. If you're excited about this role and believe you could do it well, we encourage you to apply even if your experience doesn't match every qualification listed - you may be exactly who we're looking for.

Equal Opportunity

Niantic Spatial is an equal opportunity employer. Individuals seeking employment at Niantic Spatial are considered without regard to race, color, ancestry, national origin, religion, creed, age, gender (including pregnancy, childbirth, breastfeeding or related medical conditions), marital status, physical or mental disability, medical condition, genetic information, military or veteran status, gender identity, gender expression, sexual orientation, or any other protected category under applicable laws. Niantic Spatial will also consider qualified applicants with criminal histories in accordance with applicable laws. Please contact your recruiter if you want to request an accommodation for the job application or interview process.

Candidate Privacy

I understand that by submitting my job application, the information I provide as part of that application will be used in accordance with Niantic Spatial's Privacy Notice for Job Applicants and Candidates https://www.nianticspatial.com/applicant-privacy-notice.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Niantic's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Niantic's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Niantic's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.