Skip to content

Open nowPosted 7 hours ago

Staff Engineer, Lustre

DDN99 open roles

Pay
$200,000 – $250,000 a year
Where
Santa Clara Office
Work mode
Hybrid
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Engineer, LustreDDN · Santa Clara Office
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on DDN's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market. DDN postings stay open a median of 21 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.7%30 days
This job: posted 7 hours ago

DDN median: 21 days open

The posting

We are seeking a Staff Engineer with 10+ years of experience in distributed storage and Linux-based systems engineering. This is a hands-on senior technical role focused on design, debugging, performance, and operational excellence across LustreFS and adjacent stack components. The ideal candidate brings strong expertise in one or more Lustre subsystems, can independently drive complex investigations, and collaborates effectively across engineering, QE, support and release teams. Engineers who are comfortable using AI to accelerate triage, debugging, code comprehension and new feature design will be especially valuable.

KEY RESPONSIBILITIES

- Design, develop and debug LustreFS features, fixes and enhancements across relevant subsystems such as llite, MDS/MDT, OSS/OST, LDLM and LNet.

- Investigate customer and scale-related defects, drive root-cause analysis and implement high-quality fixes with strong attention to correctness and maintainability.

- Contribute to performance tuning, failure analysis and reliability improvements for large-scale Lustre deployments.

- Participate actively in code reviews, design reviews and subsystem discussions, bringing rigor to testing and operational readiness.

- Work closely with QE and support to reproduce issues, improve diagnostic data quality and increase coverage for high-risk failure scenarios.

- Help document subsystem behavior, debugging approaches, known failure patterns and operational best practices.

- Use AI-assisted tools where appropriate to speed up issue triage, summarize logs, improve code understanding and capture reusable lessons learned.

REQUIRED QUALIFICATIONS

- 10+ years of experience in systems software, distributed systems, storage, Linux kernel or filesystem engineering.

- Strong experience in LustreFS development, support or performance engineering with depth in at least one major subsystem.

- Strong C programming and Linux systems debugging skills.

- Working knowledge of Linux kernel internals, filesystem semantics, networking and performance analysis.

- Experience with LNet and/or high-performance transports such as RDMA, InfiniBand, RoCE or TCP-based storage networking.

- Ability to debug and resolve issues spanning multiple layers including client, server, network and backend storage.

- Strong collaboration skills and the ability to work across functions in a fast-moving engineering environment.

PREFERRED SKILLS

- Experience in HPC, AI infrastructure or large-scale parallel storage environments.

- Exposure to metadata-heavy and throughput-heavy workload characterization and tuning.

- Familiarity with ZFS, ldiskfs, NVMe-backed storage and related observability / performance tooling.

- Experience creating test plans, reproducer frameworks, runbooks or diagnostic automation.

- Comfort using AI tools to accelerate debugging, code reviews, triage, documentation and early-stage design ideation.

- Experience mentoring junior engineers or leading focused technical efforts within a subsystem.

WHAT YOU WILL WORK ON

- Hands-on development and debugging of LustreFS defects, performance issues and subsystem enhancements.

- Customer-facing and scale-related issue investigation across llite, metadata, object storage, LNet and transport layers.

- Collaborative design and implementation of reliability, observability and serviceability improvements.

- Reviewing and validating fixes through targeted tests, failure injection, log analysis and performance characterization.

- Using AI-assisted workflows to accelerate triage, debug loops, code understanding and documentation quality.

- Contributing to team redundancy by strengthening documentation, code review quality and subsystem knowledge sharing.

WHY THIS ROLE MATTERS

This role is central to building durable engineering redundancy in LustreFS: expanding deep subsystem ownership, reducing concentration risk, and accelerating next-generation delivery through strong engineering fundamentals and AI-enabled execution.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against DDN's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on DDN's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    DDN's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.