Skip to content

Open nowPosted 234 days ago

Senior AI Research Engineer

Workable (global search)108,016 open roles

Where
Islamabad, Islamabad Capital Territory, Pakistan
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior AI Research EngineerWorkable (global search) · Islamabad, Islamabad Capital Territory, Pakistan
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Workable (global search)'s own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. Workable (global search) postings stay open a median of 7 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.0%30 days
This job: posted 234 days ago

Workable (global search) median: 7 days open

The posting

About Us

At SkyLabs AI Inc., we are at the forefront of the artificial intelligence revolution. As a US-headquartered company, we conduct applied research on AI for intelligent reasoning. We specialize in complex neurosymbolic AI to solve intricate problems within software engineering. Our team is composed of world-class researchers and engineers dedicated to building the platforms and intelligent agents that will power the next generation of software. If you are passionate about building truly intelligent systems and want to make a lasting impact, join us.

The Role

We are seeking an exceptional Senior AI Research Engineer with a strong focus on training and improving LLMs across the full lifecycle—from Domain‑Adaptive Pretraining to Supervised Fine‑Tuning (SFT), RL/RLVR, and advanced post‑training techniques (e.g., reward modeling, preference optimization, RL‑VR style workflows). You will lead hands‑on training efforts, contribute to research direction, and build robust pipelines for data, training, evaluation, and deployment.

You should have in‑depth understanding of LLM internals (attention/MLP dynamics, normalization, optimization behavior, scaling laws), and modern architectures including Mixture of Experts (MoE) and other frontier design choices. You will work closely with product and platform teams to translate research ideas into performant systems.

Requirements

Key Responsibilities

  • Lead the team of AI Engineers/Researchers
  • Train and iterate on LLMs end-to-end across DAPT/CPT, SFT, and RL-based post-training (preference optimization, reward modeling, verifiable rewards, policy optimization, and related variants).
  • Design training recipes: tokenization strategy, curriculum/sampling, batch/sequence packing, optimizer + scheduler choices, stability techniques, and hyperparameter search.
  • Architect and implement efficient training systems using distributed training (data/tensor/pipeline parallelism), FSDP/ZeRO, activation checkpointing, mixed precision, and throughput optimization.
  • Develop and maintain LLM data pipelines: large-scale data ingestion, filtering, deduplication, contamination checks, domain balancing, safety filtering, and dataset versioning.
  • Perform deep EDA and root-cause analysis for training/eval regressions (loss spikes, instability, alignment drift, memorization, toxicity, domain overfitting), and implement mitigations.
  • Build evaluation and benchmarking suites: offline metrics, human preference pipelines, automatic evals, domain-specific test sets, adversarial probing, and model behavior tracking over time.
  • Apply and advance model compression and efficiency: distillation from LLMs, quantization-aware considerations, pruning/sparsity (including MoE routing behavior), and inference-time optimization tradeoffs.
  • Implement and operationalize LLMOps: experiment tracking, reproducible runs, model registry, training telemetry, artifact management, CI for training configs, and safe rollout strategies.
  • Leverage AI coding tools (e.g., Cursor, Copilot) to accelerate development while maintaining strong engineering rigor, testability, and code review standards.
  • Collaborate cross-functionally with platform, infra, and product teams to integrate trained models into agentic/production pipelines (retrieval, tool use, evaluation gates, monitoring).
  • Stay current with the field and translate new ideas (architectures, post-training, data strategies, evals) into internal experiments and improvements.

Qualifications & Skills

  • 5+ years in ML engineering / applied research, with significant hands-on experience training large models, especially LLMs.
  • Strong practical knowledge of transformer internals and modern variations (MoE, routing, load balancing, attention variants, normalization choices, long-context strategies).
  • Demonstrated experience with DAPT/CPT, SFT, and RL-based post-training (e.g., reward modeling, preference datasets, verifiable rewards, policy optimization). Ability to reason about why a method works and when it fails.
  • Excellent understanding of data science fundamentals: EDA, dataset design, bias/contamination, sampling strategies, and error analysis.
  • Proven ability to root-cause and debug training instabilities and performance regressions (numerics, data issues, infra bottlenecks, config drift).
  • Expert-level programming in Python, with strong software engineering fundamentals (data structures, algorithms, systems thinking).
  • Experience with key tooling such as PyTorch, distributed training frameworks (FSDP/DeepSpeed/ZeRO), and training orchestration.
  • Familiarity with LLMOps: tracking (e.g., W&B/MLflow), dataset versioning, reproducibility, deployment constraints, and monitoring of model behavior.
  • Comfortable building agentic pipelines around LLMs (evaluation agents, data generation/labeling loops, model-in-the-loop distillation, automated red-teaming).
  • Bonus: publications, open-source contributions, or demonstrated leadership in shipping LLM training improvements.

Who You Are

You’re a pragmatic researcher-engineer: you can read papers, run experiments, and also build the training and data systems that make results reproducible and shippable. You’re meticulous about data and measurement, fast at debugging, and comfortable making tradeoffs between quality, cost, and timeline. You communicate clearly, mentor others, and thrive in an environment where you’re trusted to drive big initiatives.

Benefits

  • Salaries in USD (income tax exemptions)
  • Work in Pakistan Timezone
  • Comprehensive health allowance
  • Monthly team events and activities
  • Relocation allowance (if you're moving to Islamabad)
  • Opportunity to work with top minds in the industry and academia
  • Startup culture where ideas are heard at all levels
  • Adequate annual, sick, casual and parental leaves
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Workable (global search)'s own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Workable (global search)'s form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Workable (global search)'s answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.