Skip to content

Open nowPosted 9 hours ago

Principal Data Scientist, Agentic AI Technical Lead

Robots and Pencils32 open roles

Where
CA Remote
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowPrincipal Data Scientist, Agentic AI Technical LeadRobots and Pencils · CA Remote
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Robots and Pencils's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.2% of postings close within 7 days. Measured by our own scanner across the market. Robots and Pencils postings stay open a median of 8 days.

Share of postings closed within
  1. 1.8%1 day
  2. 3.6%3 days
  3. 8.2%7 days
  4. 15.2%14 days
  5. 34.0%30 days
This job: posted 9 hours ago

Robots and Pencils median: 8 days open

The posting

We're looking for a Principal Data Scientist to serve as the technical lead across an entire agentic AI program. This role is ideal for a deeply experienced data science and AI leader who can set direction across many workstreams at once: owning the technical vision, guiding large teams through complex design and delivery challenges, and keeping the full AI/ML program coherent, high quality, and moving at pace.

In this role, you will be the senior technical authority over every agentic AI workstream in the program, spanning a large delivery organization. You'll set architecture and evaluation standards, manage technical delivery across streams, and act as the trusted voice on technical decisions with client leadership. You'll partner closely with executives to define the AI roadmap and make high-stakes decisions that determine how agentic AI scales across the organization.

Why This Role Matters

At Robots & Pencils, we design AI systems for a human world. Our name says it all. Robots and pencils means engineering paired with creativity, because every agent we ship has to work for real people in real workflows. That balance is baked into how we operate.

Every role here contributes directly to that mission. Here, you shape how AI systems integrate into enterprise operations, how teams move at real velocity, and how products create measurable impact for clients and the people they serve. We ship production-ready AI in 30 to 45 days. That pace demands people who take ownership, lead with craft, and care deeply about what they put their name on.

What You'll Do

Craft & Delivery

  • Provide technical oversight of all agentic AI workstreams, from problem framing and design through evaluation, deployment, and ongoing operation
  • Define the technical strategy and reference architecture for agentic systems, including multi-agent orchestration, tool and function calling, RAG patterns, vector databases, embeddings, and streaming responses
  • Manage technical delivery across streams, identifying cross-team dependencies, resolving technical blockers, and keeping quality and velocity high as the program scales
  • Set the standard for model development, experimentation, and the path from research to production across the program
  • Establish rigorous evaluation frameworks for LLM and agent performance, including quality, safety, hallucination, cost, and latency, so progress is measured and not assumed
  • Guide the design of scalable ML platforms, pipelines, and workflow orchestration that support event-driven, asynchronous operations at scale
  • Ensure AI reliability, security, and scalability across deployed systems, including observability, monitoring, and debugging in production
  • Bring an AI-forward mindset to your daily work, using tools like Claude, Cursor, and other modern AI assistants to ship higher-quality work at pace

Collaboration & Communication

  • Serve as the technical face of the AI/ML program to senior client stakeholders, building trust through clear, honest communication on progress, risk, and tradeoffs
  • Co-define the AI roadmap with executive leadership, operating as a peer in strategic technical conversations
  • Translate complex technical concepts for executive, engineering, and business audiences, turning depth into decisions others can act on
  • Align data science, engineering, product, and business teams so every workstream ladders up to shared priorities and measurable outcomes

Leadership & Influence

  • Lead and influence a large, multi-team delivery organization, setting expectations for technical excellence across every stream
  • Define and champion data science and AI engineering standards that shape how the program builds, evaluates, and operates AI systems
  • Review and elevate the quality of work across teams, giving direct feedback that makes the work better
  • Mentor technical leads and senior practitioners, developing the next generation of technical leaders
  • Own high-stakes technical decisions with program-wide weight, balancing innovation, risk, and delivery speed
  • Drive technical vision, defining not just what gets built, but how agentic AI practice evolves across the program over time

What You'll Bring

  • 12+ years of experience in data science and AI/ML, with a record of taking AI systems from research to production at enterprise scale
  • Proven track record leading technical delivery across multiple concurrent workstreams or teams, with influence that extends well beyond the immediate team
  • Experience leading, or technically overseeing, large multi-team programs, including managing technical risk, dependencies, and delivery quality
  • Exceptional stakeholder management and communication skills, including credibility with senior executives and client leadership
  • Hands-on experience with LLM and agentic systems, including prompt engineering, function/tool calling, multi-agent orchestration, RAG architectures, vector databases, embeddings, and streaming LLM responses
  • Deep expertise in model evaluation, experimentation design, and applied statistics, including evaluation approaches specific to generative and agentic systems
  • Strong proficiency in Python and the modern data science and ML stack
  • Expertise in MLOps and AI infrastructure, including model versioning, monitoring, deployment automation, and reproducibility
  • Strong software engineering fundamentals, including system design, API design, code quality, and strong unit testing practices
  • Experience with distributed systems, event-driven architectures, and workflow orchestration tools
  • In-depth experience with AWS, especially the AWS GenAI offering (e.g., Amazon Bedrock); working knowledge of other cloud platforms
  • Familiarity with both SQL and NoSQL databases, including scalable design patterns
  • Working knowledge of AI governance, responsible AI, and compliance considerations in production environments

Helpful Extras and Unique Skills

  • Experience supporting AI programs in regulated industries such as life sciences, healthcare, or financial services
  • Experience in a consulting or client-facing technical leadership role
  • Background in fine-tuning, evaluation automation, or agent safety and guardrails

You'll Do Well Here if You Are

  • A doer. You see something broken and fix it. You'd rather move on clarity than wait for certainty.
  • A fast learner who knows you don't know everything. The AI landscape changes weekly. You're senior enough to know better and curious enough to keep learning anyway.
  • Direct in a way that makes the work better. You give honest feedback. You'd rather have the hard conversation than blow smoke.
  • Obsessed with craft. You know genius is in the details. You ship exceptional, not perfect, and you don't put your name on work you wouldn't stand behind.
  • Built for ownership. You honor commitments, admit mistakes fast, and back your teammates when a decision costs something. No handoffs, no finger-pointing.
  • All in. You treat clients' businesses like your own. You take the work seriously without taking yourself seriously.
  • Resourceful when the budget, timeline, or team is tight. Constraints don't slow you down. They sharpen you.
  • Glad to be in the room with people who care as much as you do. Our teams average fifteen-plus years of experience. We hire people who push each other to do better work.
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Robots and Pencils's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Robots and Pencils's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Robots and Pencils's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.