Skip to content

Open nowPosted 140 days ago

Senior/Staff Software Engineer, Distributed Systems

Hedra9 open roles

Pay
$175,000 – $275,000 a year
Where
San Francisco
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior/Staff Software Engineer, Distributed SystemsHedra · San Francisco
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Hedra's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.2% of postings close within 7 days. Measured by our own scanner across the market. Hedra postings stay open a median of 39 days.

Share of postings closed within
  1. 1.8%1 day
  2. 3.8%3 days
  3. 8.2%7 days
  4. 15.2%14 days
  5. 34.2%30 days
This job: posted 140 days ago

Hedra median: 39 days open

The posting

ABOUT HEDRA

Hedra is the platform, models, and infrastructure for visual intelligence.

We build the systems that make large-scale visual inference fast, reliable, and accessible to developers. Our work spans model serving, compute infrastructure, scheduling and routing, APIs, and the developer platform that sits on top of it.

We’re a small, highly technical team in San Francisco, backed by a16z and other leading investors. Engineers at Hedra work across boundaries, own systems end to end, and have significant influence over both what we build and how we build it.

THE ROLE

We’re looking for a Senior or Staff Software Engineer with deep experience building and operating distributed production systems.

You’ll work on the infrastructure underlying Hedra’s inference platform: systems that schedule and route compute, serve models efficiently, handle high-throughput workloads, and remain reliable as both traffic and the number of models we support grow.

The problems are often ambiguous and don’t have obvious answers. We’re looking for someone who can reason from first principles, identify bottlenecks and failure modes before they become problems, and make thoughtful tradeoffs across performance, reliability, complexity, and cost.

You do not need to come from an AI company or already be an expert in model inference. We care much more about depth in distributed systems and your ability to apply that experience to a new problem space.

WHAT YOU’LL DO

- Design, build, and operate distributed systems that power Hedra’s inference infrastructure.

- Build systems for scheduling, routing, and managing compute-intensive workloads across heterogeneous resources.

- Improve throughput, latency, reliability, and resource utilization across our serving infrastructure.

- Design systems that remain predictable and resilient under load, partial failures, changing capacity, and unpredictable workloads.

- Own production systems end to end, including deployment, CI/CD, testing and validation, observability, monitoring, alerting, debugging, and incident response.

- Identify architectural bottlenecks and failure modes and drive solutions rather than waiting for problems to be fully specified.

- Work across infrastructure, model serving, APIs, and developer-facing systems as the platform evolves.

- Use modern agentic coding workflows as part of how you design, build, debug, and ship software.

- Help shape the technical direction of a small engineering organization where individual engineers have substantial ownership.

WHAT WE’RE LOOKING FOR

- 5+ years of software engineering experience, with significant experience building distributed backend or infrastructure systems.

- A track record of designing and operating high-availability production services at meaningful scale.

- Strong distributed systems fundamentals, including experience reasoning about concurrency, queues, retries, failure recovery, consistency, backpressure, capacity, and system behavior under load.

- Experience debugging complex production systems across multiple layers of the stack.

- Strong judgment around architectural tradeoffs, particularly reliability, performance, complexity, and operational cost.

- Experience with CI/CD, comprehensive testing and validation strategies, monitoring, alerting, and production operations.

- Strong programming fundamentals and experience working in systems-oriented backend languages. Python experience is helpful given our stack.

- Experience using agentic coding tools and workflows beyond basic prompting.

- Comfort operating in an environment where problems are often underspecified and engineers are expected to independently determine the right approach.

- Ability to communicate technical decisions clearly and collaborate with engineers across infrastructure, research, and product.

NICE TO HAVE

- Experience with inference or model-serving infrastructure.

- GPU or accelerator infrastructure.

- Compute scheduling, orchestration, or resource management.

- High-throughput or low-latency systems.

- Kubernetes or other cluster orchestration systems.

- Performance optimization and profiling.

- Developer infrastructure, APIs, SDKs, or platform engineering.

- Experience operating infrastructure across cloud and/or bare-metal environments.

BENEFITS

- Competitive compensation and equity

- 401k

- Healthcare (Silver PPO Medical, Vision, Dental)

- Lunch and snacks at the office

This role is based in San Francisco, and we work together in person five days a week.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Hedra's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Hedra's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Hedra's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.