Skip to content

Open nowPosted 2 days ago

Senior AI Engineer / Agentic AI Architect

CodeNinja9 open roles

Where
Riyadh
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior AI Engineer / Agentic AI ArchitectCodeNinja · Riyadh
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on CodeNinja's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: posted 2 days ago

The posting

Description

About CodeNinja

CodeNinja is a global software and AI infrastructure company delivering full-stack technology solutions across AI, software engineering, data, and digital transformation.

With operations across Saudi Arabia and global technology hubs, CodeNinja works with organizations across multiple industries to deliver technology solutions that support business transformation and innovation.

Our teams work across areas including AI, software engineering, data and analytics, cloud, enterprise technology, and digital transformation.

CodeNinja is looking for an experienced Senior AI Engineer / Agentic AI Architect to design, build, and deploy enterprise-scale AI solutions for the banking and financial services sector.

About the Role

In this role, you will deliver production-grade AI applications with a focus on scalability, security, observability, governance, and performance. You will work across Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and Agentic AI.

The ideal candidate will have strong software engineering fundamentals and deep expertise in LLMs, RAG, Agentic AI, and modern AI engineering practices. They will have 8–12+ years in software engineering, including 5+ years of hands-on experience in Artificial Intelligence, Machine Learning, and Generative AI.

Key Responsibilities

  • Design and develop enterprise-grade AI solutions using modern LLMs and Agentic AI frameworks.
  • Architect multi-agent systems capable of planning, reasoning, tool usage, and workflow orchestration.
  • Build production-ready RAG platforms integrating structured and unstructured enterprise data.
  • Design scalable APIs and AI services using Python and modern backend frameworks.
  • Implement robust evaluation frameworks for LLM quality, safety, and performance.
  • Optimize AI applications for latency, throughput, and infrastructure cost.
  • Deploy and manage open-source LLMs in production environments.
  • Collaborate with architects, product owners, business analysts, and DevOps teams to deliver enterprise AI platforms.
  • Ensure compliance with enterprise security, governance, and responsible AI practices.
  • Mentor engineering teams and contribute to AI best practices and reusable frameworks.

Expected Deliverables

The selected candidate should be capable of independently designing and delivering:

  • Enterprise AI platforms
  • Multi-agent AI systems
  • RAG-based knowledge assistants
  • AI copilots
  • LLM evaluation frameworks
  • Production-ready AI APIs
  • AI observability and monitoring solutions
  • Secure, scalable, and cost-optimized AI deployments suitable for enterprise production environments.

Requirements

Requirements

Required Qualifications & Skills

Experience

  • 8–12+ years of experience in Software Engineering.
  • 5+ years of hands-on experience in Artificial Intelligence, Machine Learning, and Generative AI.

AI / Machine Learning

  • Strong understanding of:
  • Machine Learning
  • Deep Learning
  • NLP
  • Transformer architectures
  • Large Language Models (LLMs)
  • Embedding models

Software Engineering

  • Expert-level Python programming.
  • Strong software engineering fundamentals.
  • Experience building production-grade backend systems.
  • RESTful API and microservices development.
  • Async programming and scalable architectures.
  • Experience with FastAPI, Flask, or similar frameworks.

Agentic AI

  • Hands-on experience designing and implementing Agentic AI solutions using one or more of:
  • LangGraph
  • CrewAI
  • OpenAI Agents SDK
  • AutoGen
  • Semantic Kernel
  • LlamaIndex Workflows
  • Experience in:
  • Multi-agent orchestration
  • Planning agents
  • Tool calling
  • Human-in-the-loop workflows
  • Memory management
  • State management
  • Agent collaboration patterns

Retrieval-Augmented Generation (RAG)

  • Strong experience building enterprise RAG platforms.
  • Embedding models, such as:
  • OpenAI
  • Voyage AI
  • BGE
  • E5
  • Instructor
  • Cohere
  • Vector databases, such as:
  • Pinecone
  • Qdrant
  • Milvus
  • Weaviate
  • ChromaDB
  • Graph databases, such as:
  • Neo4j
  • Amazon Neptune
  • Memgraph
  • Search technologies, including:
  • Hybrid Search
  • BM25
  • Dense Retrieval
  • Sparse Retrieval
  • Semantic Search
  • Metadata Filtering
  • Re-ranking
  • Knowledge Graph integration

LLM Evaluation

  • Experience designing systematic evaluation frameworks using tools such as:
  • Ragas
  • TruLens
  • DeepEval
  • OpenAI Evals
  • LangSmith Evaluation
  • Understanding of:
  • Hallucination detection
  • Faithfulness
  • Answer relevancy
  • Context precision
  • Context recall
  • Groundedness
  • Toxicity
  • Regression testing

Guardrails & Observability

  • Guardrails:
  • Guardrails AI
  • NeMo Guardrails
  • OpenAI Moderation
  • Prompt Injection Detection
  • PII masking
  • Content filtering
  • Observability:
  • LangSmith
  • Langfuse
  • Arize Phoenix
  • Weights & Biases
  • MLflow
  • Experience with:
  • Prompt tracing
  • Token analytics
  • Cost monitoring
  • Latency monitoring
  • User feedback loops
  • Production debugging

Prompt Engineering

  • Expertise in:
  • Chain-of-Thought (CoT)
  • ReAct
  • Tree of Thoughts
  • Self-Consistency
  • Few-shot prompting
  • Structured prompting
  • Function Calling
  • JSON mode
  • Prompt optimization
  • Prompt caching
  • Context window optimization
  • Token usage optimization
  • Cost optimization

Model Deployment & Inference

  • Hands-on experience deploying open-source LLMs.
  • Preferred models:
  • Llama
  • Mistral
  • Qwen
  • Gemma
  • DeepSeek
  • Inference engines:
  • vLLM
  • TensorRT-LLM
  • Ollama
  • TGI (Text Generation Inference)
  • SGLang
  • Experience with:
  • GPU optimization
  • Batch inference
  • Model serving
  • Autoscaling
  • Multi-GPU deployment
  • Quantization (GGUF, GPTQ, AWQ, FP8, INT8, INT4)

MLOps / AI Platform

  • Experience with:
  • MLflow
  • Kubeflow
  • Docker
  • Kubernetes
  • GitHub Actions / GitLab CI
  • Model versioning
  • Experiment tracking
  • Feature stores
  • Continuous evaluation
  • Continuous deployment

Cloud Platforms

  • Experience with one or more:
  • Google Cloud Platform (Vertex AI)
  • Microsoft Azure AI
  • AWS Bedrock
  • OpenAI Azure

Database Technologies

  • Experience with:
  • PostgreSQL
  • Oracle
  • MongoDB
  • Redis
  • Elasticsearch / OpenSearch

Soft Skills

  • Strong analytical and problem-solving skills.
  • Excellent communication and stakeholder management.
  • Ability to lead technical discussions and architecture reviews.
  • Experience mentoring engineering teams.
  • Ability to work in Agile delivery environments.

Banking & Industry Experience (Nice to Have)

  • Banking or Financial Services domain experience.
  • Experience with enterprise AI governance and Responsible AI frameworks.
  • Knowledge of SAMA, NCA, or other financial regulatory environments.
  • Experience building AI copilots and enterprise AI assistants.
  • Knowledge of OCR, document intelligence, and intelligent automation.
  • Experience integrating AI solutions with BPM/workflow platforms such as Appian, Camunda, or Pega.

Education & Certifications

  • Bachelor's degree in Computer Science, Artificial Intelligence, Information Technology, Engineering, or a related field.
  • The following professional certifications are an advantage:
  • Google Professional Machine Learning Engineer
  • Microsoft Azure AI Engineer Associate
  • AWS Certified Machine Learning – Specialty
  • Databricks Machine Learning Professional
  • NVIDIA AI Certifications
  • OpenAI or Anthropic ecosystem certifications (where applicable)

Key Competencies

Agentic AI | Multi-Agent Systems | Large Language Models (LLMs) | RAG | Vector & Graph Databases | LLM Evaluation | Guardrails & Observability | Prompt Engineering | LLM Deployment & Inference | MLOps | Python | APIs & Microservices | Cloud AI Platforms | Responsible AI

Benefits

Benefits

What We Offer

  • Competitive compensation based on experience and qualifications.
  • Opportunity to work on enterprise-scale technology and digital transformation projects.
  • Exposure to banking, financial services, AI, data, and emerging technology environments.
  • Professional growth and learning opportunities.
  • Collaborative and technically driven work environment.
  • Opportunity to work with experienced technology and consulting professionals.

Disclaimer

This job description is intended to convey information essential to understanding the scope of the role and is not exhaustive of all responsibilities, skills, or qualifications required. CodeNinja reserves the right to modify duties and responsibilities at any time.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against CodeNinja's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on CodeNinja's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    CodeNinja's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.