Skip to content

Open nowPosted 17 days ago

Staff Machine Learning Operations Engineer - Computer Vision

Workable (global search)108,016 open roles

Where
Woburn, MA, United States
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Machine Learning Operations Engineer - Computer VisionWorkable (global search) · Woburn, MA, United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Workable (global search)'s own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. Workable (global search) postings stay open a median of 7 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.0%30 days
This job: posted 17 days ago

Workable (global search) median: 7 days open

The posting

About ATI:

Automated Tire (ATI) is a Series-B startup revolutionizing automotive service with innovative robotic and software technology. Founded by experienced entrepreneurs and backed by major players in the automotive and tire sectors, ATI is building the next generation of tools that make tire shops and dealership service lanes faster, safer, and smarter. If you're passionate about building products that ship into real-world environments, ATI is the place for you.

Position Overview:

BrakeWise is our production brake inspection product: a mobile application paired with a camera probe that technicians use to assess pad and rotor condition during live service work. The machine learning behind it is a multi-stage pipeline of segmentation and classification models that turn raw imagery into a wear assessment a shop can act on and charge for.

That pipeline works, and it is an MVP. It runs on Cloud Functions, and it will not carry us to the customer volume we're signing. We're looking for a Staff MLOps Engineer to own it — to take it from a working prototype to a serving architecture that holds up under real throughput, with the latency, cost, and reliability characteristics a paying customer expects.

You'll own every aspect of how our models reach production and how they get better: serving infrastructure, deployment and rollback, monitoring and drift detection, the retraining loop, and the evaluation discipline that tells us whether a new model is actually an improvement. Model accuracy here has commercial consequences — a bad wear call is either a missed repair or an unnecessary one, in front of a customer.

This is also the senior cloud architecture voice on the team. You'll partner closely with our Staff Full Stack Engineer, who owns the mobile app and customer dashboard, reviewing designs and setting GCP practices across the platform rather than only within the ML stack.

Responsibilities:

  • Own the multi-stage inference pipeline (segmentors and classifiers) end to end — serving architecture, latency, throughput, reliability, and cost per inspection
  • Re-architect the pipeline off its current Cloud Functions MVP onto infrastructure that scales: containerized inference, GPU-backed or accelerated serving where it pays for itself, queueing, batching, and autoscaling
  • Own model deployment: versioning, staged rollout, canary and shadow evaluation, and fast rollback when a model regresses
  • Build and own the improvement loop — field data collection, labeling workflows, dataset versioning, evaluation harnesses, and regression suites that catch quality loss before customers do
  • Monitor model quality in production: drift detection, segmented performance analysis, and triage of real-world failures against real inspection imagery
  • Define the metrics that matter commercially — false-positive and false-negative rates on a wear call, technician override rate, unit inference cost — and report against them
  • Improve model performance directly: architecture selection, augmentation, hard-example mining, and quantization or distillation where latency and cost demand it
  • Evaluate on-device versus cloud inference trade-offs for the mobile app, and own whichever path we choose
  • Establish MLOps foundations: reproducible training, experiment tracking, CI/CD for models, and infrastructure as code
  • Serve as the cloud architecture counterpart to the Staff Full Stack Engineer — reviewing designs, setting GCP best practices, and raising the platform’s infrastructure bar
  • Work with hardware and field operations on capture quality — lighting, focus, and probe positioning — since upstream image quality sets the ceiling on model performance
  • Own production support for the ML stack, including incident response and on-call participation for inference availability
  • Proactively identify technical risks and architectural trade-offs, and communicate them clearly to leadership

Requirements

  • 8+ years of professional engineering experience, including several years owning machine learning systems in production — not solely model development
  • Demonstrated experience taking a computer vision pipeline from prototype to production scale, serving real users at meaningful volume
  • Deep experience deploying and operating segmentation and classification models, including multi-stage pipelines where one model’s output feeds the next
  • Strong cloud infrastructure background, preferably GCP — Vertex AI, Cloud Run, GKE, Cloud Functions, Cloud SQL, Pub/Sub, and Docker
  • Production-grade Python, and fluency with PyTorch or TensorFlow
  • Hands-on experience with model serving and optimization — Triton, TorchServe, ONNX, TensorRT, quantization, or equivalent
  • Experience owning deployment and support for a live system, including incident response, rollback, and on-call
  • Experience building data and labeling pipelines with dataset versioning and reproducible evaluation
  • Comfort with infrastructure as code (e.g., Terraform) and CI/CD automation (e.g., GitHub Actions)
  • Sound judgment on the accuracy, latency, and cost trade-offs that determine whether an ML product is viable
  • Excellent problem-solving, debugging, and communication skills, including with non-technical stakeholders

Preferred Qualifications:

  • On-device or edge inference experience (Core ML, TensorFlow Lite, ExecuTorch) and integration into mobile applications
  • Active learning or human-in-the-loop labeling systems
  • Computer vision on small, long-tail, or industrial inspection datasets rather than large public benchmarks
  • Experience with camera and sensor integration, or working alongside hardware teams on capture quality
  • Experience with robotics, IoT, or edge computing (ROS or similar platforms)
  • Familiarity with automotive service, dealership operations, or DMS ecosystems
  • Contributions to open-source projects

Why Join ATI:

  • Be part of a groundbreaking startup transforming automotive service technology
  • Work with a team of industry veterans and top-tier robotics and software talent
  • Own the ML platform for a product that already has paying customers — your architecture decisions set the scaling ceiling
  • Our customers are our investors, so you'll develop and test in real service lane environments
  • A genuine data advantage: proprietary inspection imagery from real shops that no public dataset can replicate
  • Clear Total Addressable Market with strong pull from B2B partners
  • Competitive salary and comprehensive benefits package
  • Prime location in Woburn, MA with on-site parking
  • Collaborative, low-ego, high-intensity work environment
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Workable (global search)'s own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Workable (global search)'s form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Workable (global search)'s answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.