Skip to content

Open nowPosted 41 days ago

AI Architect

Workable (global search)108,016 open roles

Where
Istanbul, Turkey
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowAI ArchitectWorkable (global search) · Istanbul, Turkey
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Workable (global search)'s own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. Workable (global search) postings stay open a median of 7 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.0%30 days
This job: posted 41 days ago

Workable (global search) median: 7 days open

The posting

OREDATA is a Digital Transformation & IT Consulting firm with 10+ years of proven expertise and hundreds of successfully implemented projects across the EMEA region. When you join the OREDATA team, you'll be working hand-in-hand with experts focused on tackling digital, operational, analytical & data science challenges with the greatest impact. We foster collaboration with proximity, an agile and autonomous approach and best practices and guiding principles.

We are looking for an Data Scientist (LLM) to join our team within a leading company in the aviation industry. ✈️

Read more: https://medium.com/@oredata-engineering

Apply and be part of our exciting journey!

Responsibilities

  • Design, develop, test, and productionize LLM-based and Retrieval-Augmented Generation (RAG) solutions for enterprise use cases.
  • Take an active, hands-on role in the development of internal chatbot, conversational AI, knowledge assistant, and agentic AI products, from POC/MVP through production readiness.
  • Own and contribute to the technical architecture of enterprise LLM solutions, including model selection, deployment, serving, routing, evaluation, monitoring, and integration with internal AI platforms and applications.
  • Deploy, operate, and optimize open-weight and commercial LLMs, with a particular focus on on-premise and private infrastructure. This includes taking a lead role in standing up and configuring on-premise platforms (such as Red Hat OpenShift AI) from scratch when necessary.
  • Evaluate and select appropriate models based on use-case requirements, considering quality, latency, throughput, infrastructure requirements, cost, security, licensing, and operational constraints.
  • Optimize LLM inference and infrastructure utilization through techniques such as quantization, batching, caching, model serving optimization, GPU resource management, and appropriate workload allocation.
  • Act as an advocate for AI infrastructure efficiency (AI FinOps), optimizing compute costs by balancing model performance, hardware allocation (e.g., Multi-Instance GPU), and semantic routing strategies.
  • Design and improve model routing and semantic routing mechanisms to ensure requests are handled by the most appropriate model based on use case, complexity, performance, and resource requirements.
  • Design and support agentic and tool-calling architectures, ensuring that the appropriate models, tools, and enterprise services are selected and invoked reliably and securely.
  • Contribute to the evolution of the organization’s AI Gateway and shared AI platform capabilities, including model access, authorization, quotas, routing, governance, observability, and usage controls.
  • Work with structured and unstructured data to prepare, retrieve, enrich, and optimize knowledge sources used by AI applications, including embeddings, vector search, hybrid retrieval, re-ranking, chunking, and context management.
  • Establish and improve LLM evaluation and monitoring practices, including benchmark datasets, offline and online evaluation, regression testing, output quality analysis, hallucination monitoring, and performance metrics.
  • Collaborate closely with platform, infrastructure, data, software engineering, architecture, product, and business teams to translate business requirements into scalable and operationally feasible AI solutions.
  • Provide technical direction and architectural guidance on GPU capacity, AI infrastructure utilization, model serving technologies, and platform evolution, while remaining actively involved in implementation when required.
  • Stay current with developments in Generative AI, LLMs, agentic systems, model serving, inference optimization, RAG architectures, and AI infrastructure, and evaluate their practical applicability within the enterprise environment.

Requirements

Must-Haves (Minimum Qualifications)

  • Minimum 7 years of professional experience in Artificial Intelligence, Machine Learning, Data Science, Software Engineering, Data Engineering, AI Platform Engineering, or related technical roles.
  • Strong hands-on experience with Large Language Models (LLMs) and Generative AI solutions, including experience taking AI systems beyond experimentation and into production environments.
  • Proven experience deploying, serving, operating, or optimizing LLMs, preferably in on-premise, private cloud, or enterprise containerized environments.
  • Proven experience designing and scaling AI systems for high-traffic, high-concurrency environments, ensuring latency control and graceful degradation under heavy load (e.g., handling traffic spikes).
  • Practical understanding of GPU-based LLM inference and the key factors affecting GPU memory utilization, throughput, latency, concurrency, and infrastructure efficiency.
  • Hands-on knowledge of LLM inference optimization techniques such as quantization, batching, caching, model selection, and serving optimization.
  • Experience working with open-weight models and model ecosystems/frameworks such as Hugging Face, vLLM, NVIDIA inference technologies, TGI, Triton, or comparable technologies.
  • Experience with containerized infrastructure and orchestration technologies such as Kubernetes and/or OpenShift.
  • Strong practical experience designing, building, and improving RAG-based applications, including embeddings, vector databases/search, document retrieval, chunking strategies, hybrid retrieval, re-ranking, context management, and retrieval quality optimization.
  • Experience with model routing, semantic routing, or multi-model architectures, with the ability to determine how different models should be selected and utilized.
  • Hands-on experience with agentic AI workflows, tool/function calling, orchestration patterns, and integration of LLMs with internal/external tools.
  • Strong programming skills, preferably in Python, together with solid software engineering practices including testing, API design, version control, CI/CD, code quality, and maintainable system design.
  • Strong analytical thinking and problem-solving capability, with the ability to independently investigate technical problems, evaluate alternatives, make technical decisions, and drive solutions toward production.

Nice-to-Haves (Highly Preferred)

  • Specific experience with Red Hat OpenShift AI / Red Hat AI platforms is a strong plus.
  • Experience with AI Gateway, API Gateway, model gateway, or shared enterprise AI platform architectures.
  • Good understanding of LLMOps/MLOps and model lifecycle management, including model versioning, deployment, monitoring, observability, and production governance.
  • Experience with chatbot, conversational AI, knowledge assistant, enterprise search, recommendation, intelligent automation, or similar AI-enabled products.
  • Understanding of enterprise considerations around model licensing, open-source/open-weight usage, information security, data privacy, access control, governance, and responsible AI.
  • Experience providing technical leadership, architecture guidance, design reviews, or mentoring to other engineers.

Get to know us

If you want to know more about us and what we do, then visit our website: www.oredata.com

Why Oredata?

  • Open communication, flexibility and start-up spirit
  • Learning & Development opportunities for both personal and professional growth
  • Opportunity to get company paid Professional Certificates (Google Cloud Platform, Confluent Kafka, etc)
  • Access to Online Training Platforms (Udemy, Pluralsight, A Cloud Guru, Coursera, etc.)
  • Dynamic work ecosystem where you can take initiative and responsibility
  • Opportunity to work on international projects
  • Private Health Insurance
  • Birthday Leave Policy

Kişisel verileriniz işe alım sürecinin yürütülebilmesi amacıyla veri sorumlusu sıfatıyla şirketimiz Oredata Yazılım A.Ş. tarafından işlenecektir. Kişisel verilerinizin işlenmesi ve haklarınızla ilgili detaylı bilgiye https://oredata.com/personal-data-protection-policy/ bağlantısı üzerinden ulaşabilirsiniz.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Workable (global search)'s own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Workable (global search)'s form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Workable (global search)'s answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.