Skip to content

Open nowPosted 11 hours ago

Principal, Staff Site Reliability

DigitalBridge6 open roles

Where
Boca Raton, Florida
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowPrincipal, Staff Site ReliabilityDigitalBridge · Boca Raton, Florida
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on DigitalBridge's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.8%7 days
  4. 14.3%14 days
  5. 33.7%30 days
This job: posted 11 hours ago

The posting

We are hiring a Staff SRE / Infrastructure Engineer to own the reliability, scalability, and developer experience of the platforms that power our investment, portfolio, and enterprise operations. You will set the technical direction for our hybrid cloud footprint (AWS + Azure +

on-prem), harden production for a regulated buy-side environment, and lead an emerging body of work applying AI to incident response and root-cause analysis. This is a hands-on senior IC role with material influence over architecture, tooling standards, and how engineers ship software.

What you'll do

  • Own end-to-end reliability for business-critical services — SLOs, error budgets, capacity planning, DR, and incident command — with a five-nines mindset appropriate to financial services workloads.
  • Design and evolve our multi-cloud and on-prem infrastructure across AWS, Azure, and colocated environments; drive workload placement, cost, and resilience trade-offs.
  • Build and maintain the Terraform, Ansible, and CI/CD backbone that lets product and data teams ship safely and quickly; codify golden paths and paved roads.
  • Advance observability (metrics, logs, traces, profiling) so on-call engineers can localize failures in minutes, not hours; instrument reliability as a first-class product surface.
  • Lead the buildout of AI-assisted SRE capabilities — LLM-driven triage, incident summarization, runbook synthesis, and automated RCA — with human-in-the-loop guardrails.
  • Partner with security, data platform, and application teams on hardening, patching, secrets, network segmentation, and change management appropriate to a regulated environment.
  • Participate in a leader-level on-call rotation; run blameless postmortems and drive systemic fixes to closure.
  • Mentor senior engineers; set the bar for infrastructure code review, production readiness reviews, and reliability practice across the org.

Required experience

  • 10+ years building and operating production infrastructure at scale, including hybrid cloud

+ on-prem.

  • Deep expertise in AWS and Azure compute, networking, IAM, and managed data services; comfortable with account/project topology, landing zones, and org-level guardrails.
  • Expert-level Linux, Terraform, Ansible, and modern CI/CD (GitHub Actions, GitLab CI, Argo, or equivalent).
  • Track record of measurable improvements in platform reliability, MTTR, deployment velocity, and developer productivity.
  • Strong scripting/software skills in Python and/or Go; can read and refactor application code well enough to debug across the stack.
  • Production experience with Kubernetes, service meshes, and container security posture.
  • Fluency with observability stacks (Datadog, Prometheus/Grafana, OpenTelemetry, ELK/Splunk).
  • Experience running incident command and driving durable postmortem outcomes.

Nice to have

  • Prior experience in financial services, buy-side, or another regulated environment (SOC 2, SOX, GLBA).
  • Hands-on work with AI/LLM tooling for SRE — agentic incident response, RAG over runbooks/telemetry, or automated RCA.
  • Experience with data-center automation, colocation, or modular infrastructure.
  • FinOps depth: unit economics, cost attribution, and budget governance across AWS/Azure.

At DigitalBridge, we strive to create an inclusive environment where diverse employees want to work and where they can flourish professionally. In furtherance of our culture, all qualified applicants will receive consideration for employment without regard to race, national origin, gender, age, religion, disability, sexual orientation, veteran status, marital status or any other characteristics protected by law.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against DigitalBridge's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on DigitalBridge's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    DigitalBridge's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.