Skip to content

Open nowPosted 15 days ago

ASUS AICS SG - Machine Learning Engineer

MyCareersFuture94,028 open roles

Pay
SGD 6,500 – SGD 13,000 a month
Where
East, Singapore
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowASUS AICS SG - Machine Learning EngineerMyCareersFuture · East, Singapore
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on MyCareersFuture's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.7% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.3%3 days
  3. 7.7%7 days
  4. 14.0%14 days
  5. 33.7%30 days
This job: posted 15 days ago

The posting

About ASUS

AICS is part of ASUS, a multinational company known for the world’s best motherboards, PCs, monitors, graphics cards and routers. Along with an expanding range of superior gaming, content-creation and AIoT solutions, ASUS leads the industry through cutting-edge design and innovations made to create the most ubiquitous, intelligent, heartfelt and joyful smart life for everyone. With a global workforce that includes more than 5,000 R&D professionals, ASUS is driven to become the world’s most admired innovative leading technology enterprise.

About AICS

The mission of ASUS Intelligent Cloud Services (AICS) is to build revolutionary healthcare solutions with natural language processing, computer vision, and big data analytics. We provide Software as a Service (SaaS) applications to accelerate the effective use of medical data and improve the efficiency of hospital operations, unleashing the power of data for precision healthcare and bringing transformative impact to the industry.

Job Overview

We are looking for an experienced Machine Learning Engineer (ML Ops Engineer) to build and operate the infrastructure that enables reliable, scalable, and efficient AI model deployment.

You will work closely with ML Engineers, AI Researchers, Software Engineers, and Product Teams to manage the production lifecycle of AI models, including model evaluation, release, deployment, monitoring, and updates.

The role focuses on GPU infrastructure, model serving, model lifecycle management, and AI platform reliability.

Responsibilities

GPU & Compute Resource Management

  • Design and operate infrastructure for efficient GPU resource allocation and utilization across AI workloads.
  • Manage GPU workloads in Kubernetes and containerized environments.
  • Implement resource scheduling, quotas, priorities, and workload isolation for multiple AI workloads.
  • Monitor GPU utilization, capacity, performance, and resource consumption.
  • Optimize GPU utilization and inference efficiency as workloads scale.
  • Troubleshoot GPU, container, networking, and infrastructure issues in production.

Model Lifecycle & Evaluation

  • Build and maintain processes for model versioning, evaluation, release and rollback.
  • Establish automated workflows to evaluate new model versions against defined quality, performance, and reliability criteria.
  • Design model release gates to ensure new models meet predefined requirements before production deployment.
  • Compare model versions across metrics such as model quality, latency, throughput, and resource consumption.
  • Support controlled model rollout, including canary deployment, A/B testing, and rollback.
  • Maintain model metadata, evaluation results, deployment history, and release status for traceability.

Model Serving &Deployment

  • Build and operate reliable infrastructure for self-hosted AI model inference and serving.
  • Deploy and optimize LLM inference services using vLLM or similar inference engines.
  • Optimize model serving for latency, throughput, GPU utilization, and reliability.
  • Automate model deployment and configuration changes across environments.
  • Implement deployment strategies that minimize service disruption during model updates.
  • Monitor and troubleshoot production inference workloads.

Model Gateway & AI Platform

  • Build and maintain model gateway / model routing infrastructure that provides a unified interface to multiple AI models and inference backends.
  • Support model routing, traffic management, authentication, rate limiting, and observability.
  • Enable applications and AI agents to consume models through a consistent and reliable interface.
  • Integrate different model providers and self-hosted inference services into a unified platform.

Platform Reliability &Observability

  • Build monitoring and observability for AI workloads, including:
  1. GPU utilization and health
  2. Inference latency and throughput
  3. Model quality metrics
  4. Service availability
  5. Model version and deployment status
  6. Resource consumption
  • Establish logging, metrics, tracing, alerting, and operational dashboards.
  • Investigate production incidents and perform root-cause analysis.
  • Continuously improve system reliability, scalability, and operational efficiency.

Requirements

  • 4+ years of experience in software engineering, MLOps, ML infrastructure, DevOps, or a related field.
  • Strong programming skills in Python and/or Go.
  • Hands-on experience operating production AI/ML infrastructure.
  • Strong experience with Docker and Kubernetes.
  • Solid understanding of GPU infrastructure and resource management.
  • Experience with GPU scheduling, resource allocation, monitoring, or capacity planning.
  • Experience operating model serving or inference infrastructure in production.
  • Familiarity with LLM inference and serving, preferably with hands-on experience using vLLM.
  • Experience designing or operating model evaluation and release processes.
  • Experience with CI/CD and infrastructure automation.
  • Strong understanding of Linux, networking, distributed systems, and cloud-native infrastructure.
  • Experience with monitoring and observability.
  • Strong troubleshooting and problem-solving skills.
  • Ability to collaborate effectively with ML researchers, ML engineers, and software engineers.

Nice to Have

  • Experience building or operating a Model Gateway / AI Gateway.
  • Experience with model routing, traffic management, rate limiting, and multi-model serving.
  • Experience with vLLM internals and performance tuning, such as batching, KV cache, GPU memory utilization, and concurrency.
  • Experience with Kubernetes GPU scheduling and NVIDIA GPU infrastructure.
  • Experience operating on-premises GPU clusters.
  • Experience with LLM / Generative AI / Agentic AI infrastructure.
  • Experience implementing canary deployment, A/B testing, or automated model rollback.
  • Experience building internal AI / ML platforms used by multiple teams.
  • Experience with Infrastructure as Code such as Terraform.
  • Experience working in healthcare or other regulated environments is a plus.
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against MyCareersFuture's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on MyCareersFuture's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    MyCareersFuture's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.