Skip to content

Open nowFirst seen 4 hours ago

Sr./Staff Forward Deployed Engineer

Groq3 open roles

Where
United States
Work mode
Hybrid
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSr./Staff Forward Deployed EngineerGroq · United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Groq's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.8% of postings close within 7 days. Measured by our own scanner across the market. Groq postings stay open a median of 21 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.5%3 days
  3. 7.8%7 days
  4. 14.6%14 days
  5. 34.1%30 days
This job: first seen 4 hours ago

Groq median: 21 days open

The posting

Mission As a Sr./Staff Forward Deployed Engineer focused on AI Infrastructure at Groq, you will work at the frontier of large-scale AI systems, taking complex customer infrastructure programs from requirements to working production environments. You’ll help bring some of the newest accelerator and infrastructure technologies into production, spanning next-generation NVIDIA GPU systems alongside Groq’s purpose-built inference platform.

You will operate across GroqCloud, GroqMetal (Groq’s infrastructure platform), GPU and LPX infrastructure, and the networking, storage, orchestration, observability, and workload layers around them. This is an opportunity to work on infrastructure where the playbooks are still being written: bringing up new systems, solving problems that emerge only at scale, and helping customers deploy demanding AI workloads on platforms at the leading edge of the market. You will work directly with customers while partnering closely with Commercial, Field Engineering, Platform and Cloud Engineering, Networking, Data Center Operations, Security, and Support teams.

This is a deeply hands-on individual-contributor role. You will work directly in systems, write code and automation, troubleshoot across layers of the stack, and turn ambiguous customer requirements into deployed and validated solutions. The work you do in the field will also shape what comes next: turning hard-won lessons into reusable tooling, deployment patterns, reference architectures, and improvements to the Groq platform for the customers that follow.

Location: We prioritize hiring in or near the SF Bay Area, New York City and Dallas.

Responsibilities & Opportunities in This Role Own technical execution across complex customer engagements, from discovery and architecture through PoCs, demos, deployment, cluster bring-up, validation, acceptance, production readiness, and operational handoff. Translate incomplete or ambiguous customer requirements into practical architectures, implementation plans, test criteria, runbooks, and concrete engineering actions. Work hands-on across Linux, bare-metal infrastructure, Kubernetes and Slurm, networking, storage, observability, automation, and Groq platform integrations to bring customer environments online and resolve issues. Support large-scale GPU and LPX deployments, including infrastructure bring-up, cluster health and performance validation, workload testing, benchmarking, failure isolation, and production-readiness evidence. Understand customer AI workloads well enough to reason about training and inference behavior, concurrency, throughput, latency, data movement, caching, scheduling, and infrastructure bottlenecks. Lead technical portions of customer discovery, architecture reviews, demonstrations, and proofs of concept, clearly explaining design choices, tradeoffs, performance results, and risks to both engineering and business stakeholders. Partner with Networking and Security teams on customer requirements such as private connectivity and peering, routing, ingress and egress, load balancing, network policy, access controls, security architecture reviews, and enterprise security diligence. Troubleshoot production and pre-production issues that cross organizational or technical boundaries, drive them to resolution, and coordinate the right internal experts without losing end-to-end ownership. Embed with Platform, Cloud, Infrastructure, or Operations teams when priority customer deployments expose gaps that require concentrated engineering execution, automation, or integration work. Build reusable tools, automation, reference architectures, test suites, deployment patterns, documentation, and lessons learned so that customer-specific engineering makes the platform better for the next deployment. Bring structured customer feedback and field evidence back to Product and Engineering, identifying recurring gaps and helping turn one-off solutions into repeatable platform capabilities.

Ideal Candidates Have/Are 4+ years of hands-on experience building, deploying, operating, or troubleshooting cloud infrastructure, AI infrastructure, HPC systems, large-scale platforms, or similarly demanding production environments. Strong Linux and distributed-systems fundamentals, with practical experience in Kubernetes, Slurm, bare-metal environments, or comparable infrastructure platforms. Meaningful technical depth in at least one area such as GPU or accelerator systems, networking, storage, orchestration/platform engineering, or infrastructure reliability, with enough breadth to troubleshoot across adjacent layers. Working knowledge of AI training and inference workloads and how workload characteristics affect compute, networking, storage, scheduling, latency, and throughput. Strong Python, Go, Bash, or equivalent scripting/programming skills for diagnostics, automation, deployment tooling, testing, or integrations. A track record of personally debugging and delivering systems rather than operating only at the architecture, project-management, or escalation level. Ability to break ambiguous problems into concrete technical actions and drive issues to resolution when responsibility spans multiple teams. Strong written and verbal communication skills, including the ability to gather requirements from customer engineers, explain technical tradeoffs clearly, and document work so that others can reproduce it. Comfortable operating in a fast-moving environment where customer requirements, platform capabilities, and implementation details can evolve in parallel.

Preferred Qualifications Experience at a neocloud, hyperscaler, AI infrastructure provider, HPC environment, frontier AI company, or other organization operating large-scale accelerator infrastructure. Hands-on experience with NVIDIA GPU infrastructure and technologies such as CUDA, NCCL, NVLink/NVSwitch, DCGM, GPU Operator, InfiniBand, RoCE, Kubernetes, or Slurm. Experience bringing up, qualifying, or operating multi-node GPU clusters, including health checks, burn-in or stress testing, collective-communication testing, performance benchmarking, and acceptance criteria. Familiarity with high-performance storage systems such as VAST, Weka, Lustre, Ceph, or similar technologies and the data-access patterns of distributed AI workloads. Experience with infrastructure automation and lifecycle tooling such as Terraform, Ansible, CI/CD, BMC/Redfish, PXE/iPXE, or related systems. Prior solutions engineering, sales engineering, solutions architecture, or technical pre-sales experience in cloud, networking, security, AI infrastructure, or data center systems. Customer-facing networking experience including private interconnects and peering, BGP and routing, load balancing, Kubernetes/Cilium network policy, and north-south and east-west traffic design. Customer-facing security experience including access-control architecture, network isolation, enterprise security reviews, and SOC 2 / ISO 27001-style diligence or questionnaires. Experience defining or executing technical PoCs, reference architectures, cluster acceptance tests, performance benchmarks, migration plans, or production-readiness criteria.

Compensation Groq is committed to providing competitive compensation through our Total Cash philosophy, which incorporates potential bonus value directly into base pay. The total cash salary ranges for this position, which is inclusive of the potential bonus value, is dependent by level: - Staff: $270,400-$318,100 - Sr. Staff: $341,400 - $401,600 Individual placement within these ranges is determined by your geographic location, experience, skills, and alignment with internal compensation standards. These ranges are specific to candidates located in the United States. Compensation for international candidates will vary based on local market dynamics. Beyond cash compensation, Groq also offers a Long-Term Incentive (LTI) Program and a robust suite of employee benefits.

US Job Posting This position may require access to technology and/or information subject to U.S. export control laws and regulations, including the Export Administration Regulations (EAR). To comply with these requirements, candidates for this role must meet certain citizenship or residency criteria. Specifically, they must qualify as U.S. Persons for export control purposes (i.e., U.S. citizen, U.S. lawful permanent resident (Green Card holder), or a protected individual under 8 U.S.C. § 1324b(a)(3) such as a refugee or asylee), or otherwise be eligible for an applicable export license. #LI-MS1

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Groq's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Groq's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Groq's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.