Skip to content

Open nowPosted today

Senior Data Engineer – AI

MyCareersFuture94,028 open roles

Pay
SGD 7,000 – SGD 12,500 a month
Where
Singapore
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior Data Engineer – AIMyCareersFuture · Singapore
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on MyCareersFuture's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.7% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.3%3 days
  3. 7.7%7 days
  4. 14.0%14 days
  5. 33.7%30 days
This job: posted today

The posting

You will operate across both fast-moving Forward Deployed Engineering (FDE) engagements (POC/POV, pilot deployments for strategic and lighthouse clients) and steady-state system development and maintenance work — bringing the same rigor and a reusable, asset-fed approach to both.

Responsibilities:

Data Pipeline Engineering & AI-Readiness

• Design and build ingestion, cleaning, and transformation pipelines that turn messy, real-world client data into AI-ready datasets.

• Build batch and streaming pipelines (Airflow/Prefect/Kafka) that keep data flowing reliably into AI systems without manual intervention.

• Own data quality — deduplication, schema validation, completeness checks — upstream of any model or RAG pipeline.

• Proactively flag data gaps or quality issues that would degrade model/RAG performance downstream, before they surface as an AI Engineer's problem in testing.

RAG & Vector Store Architecture

• Architect document/data ingestion and indexing pipelines for Retrieval-Augmented Generation (RAG) systems — chunking strategy, embeddings, hybrid/vector search.

• Design and operate vector database and search infrastructure (pgvector/Pinecone/OpenSearch) at production scale and query volume.

Data Governance &Compliance

• Implement PII redaction, data residency, and access-control patterns aligned to PDPA and sector-specific requirements (Healthcare, Government, Transport).

• Maintain clear data lineage and metadata governance so engagement teams and auditors can trace how client data flows into AI outputs.

FDE &Development/Maintenance Coverage

• During FDE engagements: rapidly assess and prepare a client's data landscape during Discover/POC, identifying data-readiness gaps early.

• During system development & maintenance engagements: build and operate production-scale data pipelines handling the full volume and complexity of live client systems (e.g., Healthcare or Transport data at scale).

• Contribute reusable ingestion/indexing patterns back into the shared internal asset library to accelerate future engagements.

Collaboration &Leadership

• Partner closely and continuously with AI Engineers and AI Architects — understanding what a given model, RAG pipeline, or agent actually needs from the data layer and translating that into concrete pipeline and schema design decisions.

• Own the definition of "AI-ready" data for each engagement jointly with AI Engineers — agreeing on chunking strategy, metadata, freshness, and quality thresholds before pipelines are built, not after retrieval quality suffers.

• Sit in solution design conversations alongside AI Engineers and AI Architects, so data architecture and model/RAG architecture are designed together rather than data being treated as a downstream dependency.

• Mentor junior data engineers and set data engineering standards across engagements.

Requirements:

• 10+years in data engineering, including production-scale pipeline design (not just analytics/reporting pipelines).

• Strong SQL and at least one systems language (Python/Scala/Java); hands-on with batch and streaming frameworks (Airflow, Spark, Kafka).

• Experience building data pipelines for AI/ML or RAG use cases — embeddings, vector indexing, hybrid search.

• Solid understanding of data governance, PII handling, and access-control patterns in regulated environments.

• Comfortable moving between fast, exploratory data assessment (FDE/POC) and disciplined, high-volume production pipeline engineering (system development &maintenance).

• Working understanding of core AI/LLM concepts — tokenization, embeddings, chunking strategy, context windows, RAG, and agentic workflows — sufficient to hold areal technical conversation with AI Engineers and AI Architects about what "AI-ready" data means for a given use case, not just how to move and clean it.

Preferred Qualifications

• Experience with vector databases (pgvector, Pinecone, Weaviate) and search platforms(OpenSearch/Azure AI Search).

• Exposure to Singapore Government data environments (GCC/HCC) and compliance regimes(IM8, PDPA).

• Experience with sector-specific data complexity — Healthcare (clinical data governance) or Transport/Aviation systems.

• Familiarity with data cataloguing and lineage tooling.

• Prior experience embedded within an AI/ML delivery team (not just a data platform team) — i.e., has sat alongside AI Engineers day-to-day and adjusted pipeline/schema design based on model or RAG performance feedback.

Tech Stack(Illustrative)

• Languages: Python, SQL (Scala/Java a plus)

• Pipelines: Airflow/Prefect, Spark, Kafka/Debezium

• Storage/Search: Postgres, S3/Blob, pgvector/Pinecone/Weaviate, OpenSearch/Azure AI Search

• Governance: Presidio (PII redaction), data catalogue/lineage tooling

• Cloud: AWS/Azure/GCP; GCC/HCC exposure a plus

Interested candidates may send their CV to MAC (Reg No. R1221300) [email protected] quoting the job title in the Subject line. We regret that only shortlisted candidates will be notified.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against MyCareersFuture's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on MyCareersFuture's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    MyCareersFuture's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.