Skip to content

Open nowPosted 10 days ago

AI Quality and Evaluation Lead

firstnational34 open roles

Where
16 York St, Toronto, ON, Canada
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowAI Quality and Evaluation Leadfirstnational · 16 York St, Toronto, ON, Canada
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on firstnational's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.7%30 days
This job: posted 10 days ago

The posting

We are hiring a AI Quality and Evaluation Lead, Application Development!

Reporting To:

Assistant Vice President IT, Application Development

Full-Time/Part- Time:

Full-time

Posting Date:

September 22, 2026 

Closing Date:

October 20, 2026 

Hours of Work:

8:30 a.m. – 5:00 p.m.

Grade: Office Location:     

14.4 Toronto Great location! Steps away from the main public transit station

What we offer:

Highly competitive compensation package which includes, base salary, bonus, benefits, and career advancement opportunities! *Eligibility for benefits is dependent on the terms of employment  

  The Opportunity: The AI Quality and Evaluation Lead is accountable for the evaluation strategy, quality standards, release gates, guardrails, and continuous monitoring used across First National AI Factory solutions.    This role defines what good enough to ship means for AI enabled mortgage workflows, and converts business policy, risk appetite, user expectations, and production evidence into measurable acceptance criteria. Working within the AI Platform Squad, leads hands on evaluation for residential underwriting use cases with the mandate to evolve a scalable quality capability as the AI Factory grows, including reusable test assets, automated continuous evaluation, consistent scorecards, and clear accountability between product squads and shared quality teams.                                                                                                                                                             How you will contribute:

Own and continuously evolve the AI Factory quality and evaluation strategy, including risk tiers, standard metrics, acceptance thresholds, representative test suites, release evidence, monitoring expectations, and continuous improvement practices across product squads. 

Define use case specific quality measures with the Product Manager, the Residential Underwriting Business SME, the AI Engineering Lead, QA Engineers, Risk, Compliance, and Security, covering policy adherence, factual grounding, completeness, precision, recall, false positives, false negatives, confidence, explainability, user impact, and appropriate escalation. 

Own the methodology for evaluation harnesses, golden datasets, adversarial tests, regression suites, human review protocols, and automated scoring workflows across document intelligence, retrieval, summarization, agent tool use, deterministic calculations, and multi step workflows, with the AI Platform Squad building and operating the underlying evaluation infrastructure. 

Establish the quality and evaluation gates within First National production readiness standards, jointly with the AI Engineering Lead who owns engineering and operational readiness, including UAT and parallel run evidence, minimum performance thresholds, guardrail validation, privacy and access control checks, rollback criteria, kill switch readiness, runbooks, and approval records. 

In collaboration with Information Security conduct failure mode testing for prompt injection, data leakage, unsupported claims, policy deviations, tool misuse, unsafe write back, inappropriate automation, model degradation, and other AI specific operational risks. 

Monitor production quality using evaluation results, model and prompt versions, drift, latency, user feedback, exception rates, override patterns, incidents, and token or platform cost, and define the triggers for investigation, rollback, retraining, prompt changes, or workflow redesign. 

Create defect taxonomies and root cause practices that distinguish model, retrieval, prompt, data, rules, integration, user experience, and process issues, and work with product and engineering teams to prioritize durable fixes. 

Evolve the function from hands on evaluation of the first use cases to a federated operating model, with reusable scorecards, automated continuous evaluation, trained squad level QA practices, oversight of AI Quality Analysts as volume grows, independent model and vendor benchmarking, and audit ready quality evidence. 

  The experience you need:

Bachelor’s degree in computer science, engineering, data science, statistics, mathematics, information systems or a related discipline. 

7 plus years of progressive experience in quality engineering, software testing, model validation, machine learning, data science, risk analytics, AI governance, or related technology disciplines. 

Experience designing evaluation frameworks and release criteria for production AI or generative AI systems, including RAG, document intelligence, agents, model and prompt lifecycle management, observability, and human in the loop controls. 

Strong analytical and technical skills with Python, SQL, test automation, data analysis, statistical methods, and evaluation tooling such as MLflow, Ragas, DeepEval, or comparable frameworks. 

Working knowledge of accuracy and retrieval metrics, confidence scoring, sampling, experiment design, regression testing, drift monitoring, adversarial testing, red teaming, guardrails, privacy, security, and operational risk controls. 

Ability to translate business policy and risk appetite into measurable acceptance criteria, and to explain quality, risk, and trade off implications clearly to technical, business, control, audit, and senior stakeholders. 

Experience building or scaling a quality capability through reusable standards, automated test assets, continuous monitoring, governance evidence, coaching, and leadership of analysts or cross functional quality communities, including exposure to model risk, credit risk, or independent validation environments. 

Experience in financial services, mortgage lending, lending operations, servicing, broker channels, third party partnerships, or regulated technology environments is preferred. 

  Relationships: External: Engages implementation partners, cloud providers, data and AI platform providers, model providers, and specialist vendors on evaluation methods, independent benchmarking, quality evidence, and knowledge transfer to First National.  Internal: Works within the AI Platform Squad while maintaining assessment independence from the delivery teams whose work it gates.     Working Environment and Physical Demands Analysis:

Office environment Periods of high volume with tight timelines Long periods of stationary position/sitting Prolonged periods of repetitive movement (i.e. using a keyboard and mouse) Long periods of time in viewing a computer screen Multi-tasking may include speaking to customers on a telephone call while looking up information on a computer program.

  Why join First National?

Competitive Compensation  Comprehensive benefits program (i.e., Health Spending Account, Maternity and Parental Leave Top Up) Extensive training programs to set our employees up for success Modern office environment conducive to collaboration Supportive teamwork culture Opportunities to give back to the communities and work through events focused on a variety of charities Ongoing social events throughout the year

  The team you’ll join: Founded in 1988, First National is one of Canada’s largest non-bank lenders. We provide residential mortgages exclusively through the mortgage broker channel and we are Canada’s largest commercial mortgage lender. First National has been consistently recognized as a great place to work and we are proud that our employee engagement feedback is higher than our industry partners.     We would like to thank all applications for their interest in this existing vacancy, but only candidates selected for an interview will be contacted.   Artificial Intelligence is not used in our recruitment or hiring process for this role. First National is proud to be an equal opportunity employer and is committed to diversity and inclusion regardless of race, color, religion, national origin, age, gender identity, physical or mental disability, sexual orientation and any other category protected by law.   First National supports requests for accommodation from applicants with disabilities; please contact Human Resources at [email protected] should you need an accommodation at any point in the recruitment process.   #FNLOON

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against firstnational's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on firstnational's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    firstnational's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.