Skip to content

Open nowPosted 5 days ago

Senior Software Developer, ML Platform & Infrastructure

Wealthsimple Technologies44 open roles

Pay
CA$152,000 – CA$189,000 a year
Where
Toronto Headquarters
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior Software Developer, ML Platform & InfrastructureWealthsimple Technologies · Toronto Headquarters
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Wealthsimple Technologies's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.1% of postings close within 7 days. Measured by our own scanner across the market. Wealthsimple Technologies postings stay open a median of 5 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.6%3 days
  3. 8.1%7 days
  4. 15.0%14 days
  5. 33.9%30 days
This job: posted 5 days ago

Wealthsimple Technologies median: 5 days open

The posting

BUILD SOMETHING PEOPLE LOVE

Wealthsimple is Canada’s leading financial innovator. The company offers a full suite of simple, sophisticated financial products across managed investing, do-it-yourself trading, cryptocurrency, tax filing, spending and saving. Wealthsimple currently serves more than 4 million Canadians and holds over $155 billion in assets under administration. The company was founded in 2014 by a team of financial experts and technology entrepreneurs, and is headquartered in Toronto, Canada.

We're proud of what we've built — and we're just getting started. Read our Culture Manual https://www.wealthsimple.com/en-ca/culture and learn more about how we work https://www.wealthsimple.com/en-ca/careers.

ML PLATFORM & INFRASTRUCTURE TEAM

The Machine Learning Infrastructure & Platform team builds the foundational architecture powering AI and GenAI initiatives across Wealthsimple. We sit at the intersection of production MLOps and cutting-edge GenAI enablement.

As our AI footprint expands rapidly, our priority is evolving our robust MLOps foundations into a scalable, high-performance LLM serving and routing platform. We build self-serve systems that allow Data Scientists and Engineers to host open-source LLMs reliably, optimize inference latencies, manage GPU infrastructure, and benchmark model performance safely in production.

THE ROLE

We are looking for an experienced MLOps or ML Platform Engineer who is excited to pivot their deep background in model orchestration, serving, and platform tooling toward solving the unique challenges of LLM inference and infrastructure.

In this role, you will bridge the gap between traditional MLOps (model lifecycles, pipeline orchestration, serving infrastructure) and modern GenAI stack requirements (vLLM, GPU cluster management, intelligent model routing, and automated Evals). You will take end-to-end ownership of setting the technical direction for operating enterprise-grade LLM systems company-wide.

IN THIS ROLE, YOU WILL HAVE THE OPPORTUNITY TO:

- Turn MLOps expertise to GenAI: Transition traditional ML lifecycle and serving patterns into state-of-the-art LLM inference engines and GPU orchestration systems.

- Build model-routing architecture: Design low-latency routing frameworks (e.g., LiteLLM integration) to dynamically direct requests across managed cloud providers (AWS Bedrock) and self-hosted open-source models.

- Provision & scale GPU infrastructure: Architect and manage high-performance GPU serving environments on Kubernetes using engines like vLLM, Ray, and Triton.

- Develop evaluation & benchmarking tooling: Build automated Evals and observability frameworks to empower engineers and data scientists to validate model quality, latency, and drift against production requirements.

- Empower self-serve ML across Wealthsimple: Partner with product engineering and data science teams to build framework-agnostic platform tooling that abstracts infrastructure complexity.

- Drive cost & performance optimization: Improve price-performance across self-hosted and managed inference by optimizing capacity, utilization, batching, routing, and model selection while meeting quality and reliability objectives

WE ARE LOOKING FOR PEOPLE WHO HAVE:

- 7+ years of software engineering experience in ML Infrastructure, MLOps, ML Tooling, or Data Platform engineering.

- Deep experience in MLOps/ML Platform practices: Proven track record building and operating self-serve ML platforms, model registry workflows, experiment tracking, or production serving infrastructure (Kubeflow, MLflow, Ray, Triton, SageMaker).

- Strong platform fundamentals: Advanced proficiency in Python, container orchestration via Kubernetes, infrastructure-as-code (Terraform), and cloud provider ecosystem (AWS).

- Strong appetite to specialize in LLM serving: A genuine desire to leverage your existing MLOps skillset to tackle LLM-specific challenges (vLLM, model routing, prompt engineering tooling, vector databases, GPU memory optimization, or LLM evaluation frameworks).

- Backend performance & observability focus: Experience designing highly available, observable microservices (e.g., FastAPI) handling real-time, low-latency requests.

- End-to-end technical ownership: Proven capability to lead architectural roadmaps, guide multi-functional projects with high autonomy, and maintain complex platform systems for the long run.

NICE-TO-HAVES (OR AREAS YOU WILL LEARN ON THE JOB):

- Direct experience serving open-source Large Language Models in production (vLLM, SGLang, TensorRT-LLM, Dynamo).

- Hands-on work with CUDA, GPU partitioning, or distributed inference frameworks (Ray Serve).

- Familiarity with vector search and retrieval engines (Elasticsearch, Qdrant, Pinecone).

OUR STACK INCLUDES:

- Container & Infrastructure: Kubernetes, Terraform, AWS GPU Infrastructure

- ML Serving & LLM Tooling: vLLM, Ray, Triton, LiteLLM, AWS Bedrock, MLflow, SageMaker

- Languages & Frameworks: Python, FastAPI, PyTorch

- Data & Streaming: Kafka, Postgres, Redshift, Snowflake

WHY WEALTHSIMPLE?

🌸 Top-tier health benefits and life insurance

📈 Long-term group savings with employer match, through Wealthsimple for Business

🌴 20 vacation days, 4 wellness days, and unlimited sick and mental health days per year*

✈️ 90 days away: work outside Canada for up to 90 days per year*

👥 Employee resource groups, including Rainbow (2SLGBTQ), Women of WS, and Black at WS

🌎 We are a hybrid team with over 1,500 employees across North America. The people are one of the best parts of working here: you'll collaborate with incredibly talented, curious, and driven teammates who are deeply committed to doing great work.

*Unlimited paid sick days, Wellness Days and the 90 day away program do not apply to certain roles.

ICYMI

Technology & Innovation at Wealthsimple: We move quickly and build thoughtfully. That means we're always looking for better ways to work — whether that's new tools, AI, or rethinking how we approach a problem. We don't expect you to have all the answers, but we do expect curiosity and a willingness to evolve alongside the products we're building.

Inclusion Statement: We're building products for a diverse world, and we need a diverse team to do it well. We strongly encourage applications from everyone, regardless of race, religion, colour, national origin, gender, sexual orientation, age, marital status, or disability status.

Accessibility Statement: We're committed to an accessible hiring experience. If you need any accommodations throughout the interview process, please let us know — we'll work with you to make sure you have what you need. We also welcome any feedback on how we can better accommodate candidates with accessibility needs.

AI in Hiring: We may use artificial intelligence (AI) tools to support parts of our hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our team but don't replace human judgment – all final hiring decisions are made by people. If you have questions about how your data is used, reach out to us.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Wealthsimple Technologies's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Wealthsimple Technologies's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Wealthsimple Technologies's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.