Skip to content

Open nowPosted 13 hours ago

Sr. Reliability Engineer

hims & hers141 open roles

Pay
$140,000 – $165,000 a year
Where
US Remote
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSr. Reliability Engineerhims & hers · US Remote
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on hims & hers's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.2% of postings close within 7 days. Measured by our own scanner across the market. hims & hers postings stay open a median of 7 days.

Share of postings closed within
  1. 1.8%1 day
  2. 3.6%3 days
  3. 8.2%7 days
  4. 15.2%14 days
  5. 34.0%30 days
This job: posted 13 hours ago

hims & hers median: 7 days open

The posting

Hims & Hers is the leading health and wellness platform, on a mission to help the world feel great through the power of better health. We are redefining healthcare by putting the customer first and delivering access to care that is affordable, accessible, and personal, from diagnosis to treatment to delivery. No two people are the same, so we provide access to personalized care designed for results. By normalizing health & wellness challenges and innovating on their solutions, we’re making better health outcomes easier to achieve.

Hims & Hers is a public company, traded on the NYSE under the ticker symbol “HIMS.” To learn more about the brand and offerings, you can visit hims.com/about http://hims.com/about and hims.com/how-it-works http://hims.com/how-it-works . For information on the company’s outstanding benefits, culture, and its talent-first flexible/remote work approach, see below and visit www.hims.com/careers-professionals http://www.hims.com/careers-professionals.

ABOUT THE ROLE:

We're looking for a Senior Reliability Engineer to make Hims & Hers systems measurably more reliable: building the observability, tooling, and automation that catch problems before people do. Recent work at this level includes preparing our stack for 32x baseline traffic during high-stakes seasonal surges with zero customer-facing issues and improved P95 latency, turning databases into a paved platform with consistent observability and guardrails, and building AI agents that pull FireHydrant, Datadog, Jira, and Confluence into a single incident picture. We use AI first: every engineer gets a Claude Enterprise license, and we expect you to use it as a core part of how you investigate, build, and write, not as an afterthought.

YOU WILL:

- Own reliability for Tier 1 customer journeys: Define and instrument SLOs, golden signals, and business-level monitors for the journeys that matter most (checkout, telehealth visits, prescription fulfillment), so that degradation is detected by systems rather than by customers or support tickets. Move our "detected by monitors vs. detected by humans" ratio in the right direction and be able to show it.

- Engineer for peak load and failure: Lead capacity and resilience work for high-stakes events and steady-state growth: load-test design, second-by-second analysis of prior events, database and backend bottleneck investigation, and hardening across caching, GraphQL, VPC capacity, and vendor rate limits. Validate before game day, not during it.

- Build incident response as software: Mature FireHydrant, Datadog, and Jira into one connected pipeline: automated incident and RCA ticket creation, SLO burn-rate and composite alerting with deduplication, Tier 1 alert routing, and enforced post-mortem action tracking. Author and maintain the runbooks and severity standards that make any responder effective on any service.

- Automate operational excellence with AI: Build and operate agents and tooling that reduce manual OE work: OER report generation, RCA drafting, stale action-item detection, monitor and runbook gap detection, and OpenClaw agents wired to Datadog and FireHydrant for first-pass incident triage. Ship these as reusable capabilities, not personal scripts.

- Be the deep-debugging expert, and make teams better at it: Teams own debugging their own services, but you are the person they pull in when a problem crosses boundaries: silent service-to-service failures, anomalous traffic, or regressions that span frontend, API, mesh, and database layers. Bring that depth to the hardest cases yourself, then turn what you learn into runbooks, tooling, and pairing so teams can catch and resolve the next one on their own. Partner with Security and product teams on anomaly detection and response, weighing engineering cost and user impact alongside the benefit of any control.

- Maintain the tooling that informs Operational Excellence reviews: Produce the tooling that powers the metrics and narrative that go into bi-weekly VP-level OE reviews and the monthly cross-engineering OER, and drive the resulting action items to closure.

- Raise the bar for others: Document what you build, onboard teammates to roll it out, and coach engineers across squads on SLOs, blameless post-mortems, and on-call practice.

You Have:

- 5+ years as a Software, SRE, Platform, or Infrastructure Engineer, with a track record of owning reliability outcomes for production systems that customers depend on.

- Strong software engineering fundamentals. You solve reliability problems by writing code and building tooling, and you're comfortable reading application code across the stack to find the real cause.

- Hands-on depth in observability and SLO engineering: golden signals, burn-rate alerting, journey-level monitors, and turning noisy alert streams into actionable pages (Datadog preferred; Prometheus/Grafana and OpenTelemetry welcome).

- Production experience with AWS, Kubernetes/EKS, Terraform, and PostgreSQL (RDS/Aurora).

- Experience running or maturing incident management end to end: on-call design, escalation policies, incident command, blameless post-mortems, and action-item follow-through in a tool like FireHydrant or PagerDuty.

- Daily, practical use of AI coding and analysis tools (Claude, Cursor, or similar) to accelerate investigation, code, documentation, and reporting, and clear judgment about when to trust the output and when to verify it.

- Communication skills to explain risk, tradeoffs, and post-incident learnings to engineers and leadership alike, in writing and in review meetings.

PREFERRED QUALIFICATIONS:

- Experience building AI agents or LLM-backed automation for operations: incident triage, RCA generation, alert correlation, or observability data analysis, ideally with MCP or tool-calling integrations against Datadog, FireHydrant, or Jira.

- Load-testing and performance-engineering experience at meaningful scale (k6 or similar), including validating systems to well above expected peak.

- Familiarity with service mesh (Istio) and its observability and traffic-management features.

- Background in a regulated or healthcare environment, where reliability and data handling carry patient-safety and compliance weight.

- Experience designing vendor and partner escalation frameworks with defined severities and response SLAs.

OUR BENEFITS (THERE ARE MORE BUT HERE ARE SOME HIGHLIGHTS):

- Competitive salary & equity compensation for full-time roles

- Unlimited PTO, company holidays, and quarterly mental health days

- Comprehensive health benefits including medical, dental & vision, and parental leave

- Employee Stock Purchase Program (ESPP)

- 401k benefits with employer matching contribution

- Offsite team retreats

We are committed to building a workforce that reflects diverse perspectives and prioritizes ethics, wellness, and a strong sense of belonging. If you're excited about this role, we encourage you to apply—even if you're not sure if your background or experience is a perfect match.

Hims considers all qualified applicants for employment, including applicants with arrest or conviction records, in accordance with the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance, the California Fair Chance Act, and any similar state or local fair chance laws.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Hims & Hers is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please contact us at [email protected] and describe the needed accommodation. Your privacy is important to us, and any information you share will only be used for the legitimate purpose of considering your request for accommodation. Hims & Hers gives consideration to all qualified applicants without regard to any protected status, including disability. Please do not send resumes to this email address.

To learn more about how we collect, use, retain, and disclose Personal Information, please visit our Global Candidate Privacy Statement https://www.hims.com/global-candidate-privacy-statement.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against hims & hers's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on hims & hers's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    hims & hers's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.