Skip to content

Open nowPosted 11 hours agoWe saw it 44 min after it went up

Site Reliability Engineering Lead

Jobgether4,394 open roles

Where
US
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSite Reliability Engineering LeadJobgether · US
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Jobgether's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market. Jobgether postings stay open a median of 4 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.7%30 days
This job: posted 11 hours ago

Jobgether median: 4 days open

The posting

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Site Reliability Engineering Lead based in the United States.

This leadership role is responsible for building reliable, scalable, and secure cloud platforms while developing a high-performing SRE team. You will lead engineers, establish priorities, and drive initiatives that improve service reliability, resilience, automation, and operational efficiency. The role combines people leadership with hands-on technical direction across modern cloud and infrastructure environments. You will work closely with Development, Security, Product, and other engineering teams to strengthen platform performance and incident response. A key focus will be reducing operational toil through automation, observability, infrastructure as code, and self-healing capabilities. You will also guide post-incident reviews, root-cause analysis, and continuous improvement efforts across services and infrastructure. This is an opportunity to shape SRE practices at scale while supporting engineers in their technical and professional growth.

Accountabilities

  • Lead, mentor, and develop a small to medium-sized team of Site Reliability Engineers through regular 1:1s, performance reviews, career planning, and ongoing coaching.
  • Own hiring, onboarding, team capacity, resourcing, and workforce planning decisions to ensure the team can effectively support business and platform priorities.
  • Establish team objectives, prioritize the engineering backlog, coordinate planning, and ensure projects and operational tasks remain aligned with reliability goals.
  • Lead reliability initiatives across infrastructure and services, improving availability, scalability, resilience, security, and operational performance.
  • Drive incident response activities and facilitate blameless post-incident reviews, ensuring timely root-cause analyses and actionable follow-up.
  • Partner with Development, Security, Product, and other engineering teams to resolve cross-functional issues and strengthen collaboration.
  • Champion automation and operational excellence by reducing manual work, eliminating recurring toil, and introducing self-healing systems and infrastructure automation.
  • Support the design and evolution of scalable, secure, cloud-native environments and continuously identify opportunities to improve performance, reliability, and cost efficiency.
  • Demonstrated experience in SRE, DevOps, infrastructure engineering, or a related discipline, including experience leading engineering teams.
  • Expert-level knowledge of Kubernetes, including cluster architecture, upgrades, autoscaling, security hardening, and large-scale troubleshooting.
  • Advanced experience with Terraform, including modular infrastructure-as-code design, state management, multi-environment provisioning, and policy-as-code.
  • Deep knowledge of Azure Cloud services, including compute, networking, identity and access management, storage, and cost optimization.
  • Experience designing and scaling CI/CD pipelines using GitHub Actions, release strategies, and automated rollback approaches.
  • Strong knowledge of observability platforms such as Prometheus, Grafana, and OpenTelemetry, along with experience managing SLOs, SLAs, and error budgets.
  • Strong automation capabilities and advanced proficiency in Python, Bash, and/or PowerShell for infrastructure tooling and operational automation.
  • Deep understanding of networking fundamentals, including TCP/IP, DNS, load balancing, VPNs, and cloud-native networking.
  • Proven experience leading incident response, conducting root-cause analysis, and implementing measurable reliability improvements.
  • Strong people leadership, communication, prioritization, and problem-solving skills, with the ability to support engineers while coordinating effectively across technical teams.
  • U.S. national base salary range of $118,300–$219,800, with geographic differentials potentially applying depending on location.
  • Eligibility for an annual incentive bonus.
  • Country- and location-specific employee benefits designed to support overall health and well-being.
  • Support for an accessible and inclusive hiring process, including reasonable accommodations where required.
  • Remote/home-based opportunities available in multiple U.S. locations, including Florida, Connecticut, New Jersey, New York, and Pennsylvania.
  • Opportunity to lead a technically sophisticated SRE function focused on cloud platforms, automation, resilience, and operational excellence.
  • Professional development opportunities through team leadership, cross-functional collaboration, and exposure to large-scale cloud environments.

How Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Jobgether's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Jobgether's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Jobgether's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.