Skip to content

Open nowPosted 11 days ago

Lead Enterprise Lakehouse Architect – Data Products & Agentic AI- Contract

MyCareersFuture94,028 open roles

Pay
SGD 10,000 – SGD 12,000 a Monthly
Where
Central, Singapore
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowLead Enterprise Lakehouse Architect – Data Products & Agentic AI- ContractMyCareersFuture · Central, Singapore
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on MyCareersFuture's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.6%30 days
This job: posted 11 days ago

The posting

Lead Enterprise Lakehouse Architect – Open Table Formats, Data Products & Agentic AI

Contract Duration: 09 months

Seniority: L4 – More than 10 years of relevant experience

Working Arrangement: Onsite ( 5 days from office )

Headcount: 1

Role Overview

We are seeking an experienced Enterprise Data Lakehouse Architect to own the end-to-end architecture of a large-scale Lakehouse platform supporting governed data products, Data-as-a-Service, real-time analytics, knowledge layers and agentic AI workloads.

This is a senior hands-on architecture position requiring demonstrable production implementation experience. Applicants whose experience is limited to traditional data warehouses, BI reporting, general cloud architecture or data-engineering delivery without end-to-end Lakehouse ownership will not meet the requirements.

Responsibilities

  • Define the technical vision, target architecture and implementation roadmap for an enterprise-scale Lakehouse platform.
  • Architect reusable, scalable and secure platform components across on-premises, hybrid and cloud environments.
  • Design and implement Bronze, Silver and Gold medallion layers using Delta Lake, Apache Iceberg or Apache Hudi.
  • Design object-storage architecture covering lifecycle management and hot, warm and cold data-tiering strategies.
  • Architect MPP and distributed-compute workloads using Spark, Databricks, BigQuery, Dataproc, EMR, Synapse or equivalent platforms.
  • Establish foundation and business data products with formal data contracts, SLAs, ownership, lineage and data-quality rules.
  • Serve governed data products to downstream applications through REST APIs, Kafka/Pub-Sub, real-time streams, dashboards and data-marketplace capabilities.
  • Design reusable patterns for structured and unstructured content ingestion, lambda processing and retrieval-augmented data workloads.
  • Enable RAG and agentic AI workloads using embeddings, vector databases, graph databases, prompt engineering and context-management strategies.
  • Design secure hybrid-cloud connectivity using private dedicated connectivity, workload-placement strategies and data-egress cost controls.
  • Implement Infrastructure-as-Code and automated platform provisioning.
  • Lead platform performance engineering, query optimisation, capacity planning, reliability improvements and FinOps initiatives.
  • Evaluate Lakehouse, federation, query-engine, vector-database and graph-database technologies through RFPs and proofs of concept.
  • Define functional, non-functional, security and solution-design specifications.
  • Review technical designs and delivery outputs for compliance with architecture, engineering, security and quality standards.
  • Integrate the Lakehouse platform with enterprise CI/CD, testing, source-control, monitoring, scheduling and incident-management tools.
  • Lead continuous service-improvement and process-improvement initiatives.

Mandatory Requirements

Applicants must meet all the following requirements:

  • Between 10 and 15 years of relevant experience in enterprise data architecture, big-data platforms and distributed data processing.
  • At least five years of hands-on architecture ownership for enterprise-scale data platforms.
  • Personally architected and implemented at least one production-scale Lakehouse in banking or financial services.
  • Hands-on implementation experience with at least one approved platform:ClouderaHuawei CloudGoogle BigQuery, BigLake, Dataplex or DataprocAWS EMR or OutpostsAzure Synapse or Azure Databricks
  • Production implementation of Bronze, Silver and Gold medallion architecture.
  • Deep hands-on experience with at least one open-table format: Delta Lake, Apache Iceberg or Apache Hudi.
  • Ability to explain ACID transactions, schema evolution, partition evolution, time travel/snapshots, compaction and small-file management.
  • Experience designing distributed Spark/PySpark workloads and performing query, storage and compute optimisation.
  • Production experience implementing both batch and real-time/streaming pipelines.
  • Hands-on Data-as-a-Service implementation using REST APIs and Kafka/Pub-Sub.
  • Experience building reusable foundation and business data products supported by data contracts, SLAs and automated data-quality controls.
  • Experience publishing governed data products through a catalogue, exchange or data marketplace.
  • Experience with enterprise object storage and hot, warm and cold lifecycle strategies.
  • Experience implementing metadata management, data lineage, RBAC, audit logging and fine-grained access controls.
  • Production experience enabling RAG workloads using embeddings and a vector database.
  • Practical knowledge of graph databases, prompt engineering, context management and LLM governance.
  • Experience designing hybrid-cloud platforms, private connectivity, workload placement and egress-cost optimisation.
  • Hands-on Infrastructure-as-Code experience using Terraform, CloudFormation or ARM/Bicep.
  • Strong CI/CD implementation experience using Jenkins, Azure DevOps, Cloud Build, GitHub Actions or equivalent.
  • Experience with platform monitoring, incident management, performance engineering and continuous service improvement.
  • Ability to work onsite at IH2, Malaysia throughout the 12-month assignment.

Mandatory Certifications

Applicants must possess at least two current professional certifications, including:

  • One professional-level cloud architecture or data-engineering certification from Google Cloud, AWS or Microsoft Azure; and
  • One Databricks Data Engineer Professional, Databricks Data Architect, CDMP or equivalent data-platform certification.

Associate-level training badges or course-completion certificates alone will not satisfy this requirement.

Preferred Experience

  • Trino, Denodo or Dremio data federation.
  • Hive, Impala or Apache Kudu query engines.
  • Migration from Teradata, Greenplum or Netezza into a modern Lakehouse.
  • Databricks Vector Search, Azure AI Search, Pinecone, Weaviate, ChromaDB or Snowflake Cortex.
  • Neo4j, JanusGraph, TigerGraph, Amazon Neptune or Stardog.
  • LangGraph, OpenAI Agents SDK, Microsoft Agent Framework, LlamaIndex Workflows or Google ADK.
  • Kubernetes or OpenShift deployment using Helm or Kustomize.
  • Banking regulatory requirements and controls covering MAS, BCBS 239, AML, data residency and auditability.

Interested candidates are kindly requested to email their CV with their experience to [email protected]

We look forward to your application!

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against MyCareersFuture's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on MyCareersFuture's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    MyCareersFuture's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.