Skip to content

Open nowPosted 7 hours ago

Storage & Virtualization Engineer

AMD1,272 open roles

Where
Austin, Texas, United States
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStorage & Virtualization EngineerAMD · Austin, Texas, United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on AMD's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.1% of postings close within 7 days. Measured by our own scanner across the market. AMD postings stay open a median of 38 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.6%3 days
  3. 8.1%7 days
  4. 15.1%14 days
  5. 34.0%30 days
This job: posted 7 hours ago

AMD median: 38 days open

The posting

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.

Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.

THE ROLE:

We are seeking a Storage & Virtualization Engineer to provide L2 technical and incident-management support for storage, virtualization, and related compute dependencies used by AMD Fleet Services. This offshore role is an escalation point for incidents that exceed L1 runbooks and is responsible for advanced diagnosis, coordinated restoration, authorized remediation, recovery validation, and complete escalation to the accountable platform or engineering owner.

The engineer will improve the reliability, availability, and operational readiness of storage and virtual infrastructure while helping the IOQ Modern NOC convert recurring issues into monitoring, automation, SOP, runbook, capacity, and resiliency improvements.

THE PERSON:

You are a hands-on infrastructure engineer with strong storage, virtualization, Linux, and systems troubleshooting fundamentals. You use evidence to isolate failures across hosts, hypervisors, virtual machines, networks, and storage paths; balance restoration speed with change discipline; and communicate ownership, impact, and risk clearly during high-priority incidents.

KEY RESPONSIBILITIES:

  • Provide L2 incident management and technical support for in-scope storage, virtualization, hypervisor, virtual-machine, and related compute services used by AMD Fleet Services.
  • Own technical investigation of assigned incidents from L1 escalation through diagnosis, mitigation, recovery validation, documentation, and handoff or closure.
  • Assess incident impact, priority, affected assets, dependencies, recovery options, risks, and required owners; provide concise updates during incident bridges and follow-the-sun handoffs.
  • Validate L1 evidence, isolate storage, virtualization, host, guest, network, operating-system, capacity, or performance failure domains, and execute approved recovery actions.
  • Investigate storage availability, latency, throughput, capacity, pathing, mount, protocol, volume, snapshot, replication, and data-access symptoms using platform telemetry and system evidence.
  • Support operational troubleshooting for storage platforms such as Weka, NetApp, Pure, Ceph, Lustre, NFS, block storage, or comparable technologies within assigned scope.
  • Troubleshoot virtualization platforms such as Xen, VMware, KVM, or comparable environments, including host health, guest state, resource allocation, datastore access, migration, and recovery validation.
  • Diagnose Linux host and virtual-machine issues involving services, filesystems, multipathing, device state, permissions, resource contention, performance, networking, and logs.
  • Correlate storage and virtualization telemetry with compute, network, identity, cluster, and application evidence to identify cross-domain dependencies and the accountable owner.
  • Coordinate restoration and escalation with storage engineering, platform engineering, compute, network, cloud, identity, security, facilities, and vendor support teams while maintaining clear IOQ scope and ownership boundaries.
  • Create complete escalation packages with impact, timeline, affected systems, evidence, actions attempted, change references, residual risk, and recommended next steps.
  • Participate in postmortems and problem-management reviews; identify recurring failure modes and track corrective actions to the accountable owner.
  • Develop and maintain SOPs, runbooks, validation checks, troubleshooting guides, and knowledge articles that enable safe and consistent L1 execution.
  • Automate repeatable diagnostics, evidence collection, capacity checks, configuration validation, health verification, and approved remediation using Python, shell, Ansible, APIs, or comparable tools.
  • Improve observability by defining actionable storage and virtualization telemetry, alerts, dashboards, service-health indicators, capacity thresholds, and diagnostic evidence requirements.
  • Maintain accurate Jira records, operational documentation, shift handoffs, and service-status updates throughout the incident lifecycle.

PREFERRED EXPERIENCE:

  • Experience supporting enterprise storage and virtualization in a production data center, GPU, HPC, private-cloud, or large-scale infrastructure environment.
  • Hands-on experience with storage platforms such as Weka, NetApp, Pure, Ceph, Lustre, NFS, block storage, or comparable technologies.
  • Hands-on experience with virtualization platforms such as Xen, VMware, KVM, or comparable technologies.
  • Strong understanding of storage protocols, multipathing, volumes, filesystems, snapshots, replication, performance, capacity, and data-protection concepts.
  • Strong Linux administration and troubleshooting across services, filesystems, devices, networking, permissions, performance, and logs.
  • Experience with server hardware, hypervisors, clustering, virtual-machine lifecycle operations, and infrastructure management tooling.
  • Automation experience with Python, shell, Ansible, REST APIs, Git, CI/CD, or configuration-management systems.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Zabbix, or comparable platforms.
  • Experience with Jira or comparable ticketing systems, major-incident response, change control, postmortems, and problem management.
  • Ability to work independently during assigned offshore coverage and provide complete handoffs to United States-based IOQ and engineering teams.

ACADEMIC CREDENTIALS:

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field preferred; equivalent relevant experience considered.
  • Storage, virtualization, Linux, cloud, automation, or IT service-management certification is beneficial.

LOCATION: Austin, TX

This role is not eligible for visa sponsorship.

#LI-TL1

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

THE ROLE:

We are seeking a Storage & Virtualization Engineer to provide L2 technical and incident-management support for storage, virtualization, and related compute dependencies used by AMD Fleet Services. This offshore role is an escalation point for incidents that exceed L1 runbooks and is responsible for advanced diagnosis, coordinated restoration, authorized remediation, recovery validation, and complete escalation to the accountable platform or engineering owner.

The engineer will improve the reliability, availability, and operational readiness of storage and virtual infrastructure while helping the IOQ Modern NOC convert recurring issues into monitoring, automation, SOP, runbook, capacity, and resiliency improvements.

THE PERSON:

You are a hands-on infrastructure engineer with strong storage, virtualization, Linux, and systems troubleshooting fundamentals. You use evidence to isolate failures across hosts, hypervisors, virtual machines, networks, and storage paths; balance restoration speed with change discipline; and communicate ownership, impact, and risk clearly during high-priority incidents.

KEY RESPONSIBILITIES:

  • Provide L2 incident management and technical support for in-scope storage, virtualization, hypervisor, virtual-machine, and related compute services used by AMD Fleet Services.
  • Own technical investigation of assigned incidents from L1 escalation through diagnosis, mitigation, recovery validation, documentation, and handoff or closure.
  • Assess incident impact, priority, affected assets, dependencies, recovery options, risks, and required owners; provide concise updates during incident bridges and follow-the-sun handoffs.
  • Validate L1 evidence, isolate storage, virtualization, host, guest, network, operating-system, capacity, or performance failure domains, and execute approved recovery actions.
  • Investigate storage availability, latency, throughput, capacity, pathing, mount, protocol, volume, snapshot, replication, and data-access symptoms using platform telemetry and system evidence.
  • Support operational troubleshooting for storage platforms such as Weka, NetApp, Pure, Ceph, Lustre, NFS, block storage, or comparable technologies within assigned scope.
  • Troubleshoot virtualization platforms such as Xen, VMware, KVM, or comparable environments, including host health, guest state, resource allocation, datastore access, migration, and recovery validation.
  • Diagnose Linux host and virtual-machine issues involving services, filesystems, multipathing, device state, permissions, resource contention, performance, networking, and logs.
  • Correlate storage and virtualization telemetry with compute, network, identity, cluster, and application evidence to identify cross-domain dependencies and the accountable owner.
  • Coordinate restoration and escalation with storage engineering, platform engineering, compute, network, cloud, identity, security, facilities, and vendor support teams while maintaining clear IOQ scope and ownership boundaries.
  • Create complete escalation packages with impact, timeline, affected systems, evidence, actions attempted, change references, residual risk, and recommended next steps.
  • Participate in postmortems and problem-management reviews; identify recurring failure modes and track corrective actions to the accountable owner.
  • Develop and maintain SOPs, runbooks, validation checks, troubleshooting guides, and knowledge articles that enable safe and consistent L1 execution.
  • Automate repeatable diagnostics, evidence collection, capacity checks, configuration validation, health verification, and approved remediation using Python, shell, Ansible, APIs, or comparable tools.
  • Improve observability by defining actionable storage and virtualization telemetry, alerts, dashboards, service-health indicators, capacity thresholds, and diagnostic evidence requirements.
  • Maintain accurate Jira records, operational documentation, shift handoffs, and service-status updates throughout the incident lifecycle.

PREFERRED EXPERIENCE:

  • Experience supporting enterprise storage and virtualization in a production data center, GPU, HPC, private-cloud, or large-scale infrastructure environment.
  • Hands-on experience with storage platforms such as Weka, NetApp, Pure, Ceph, Lustre, NFS, block storage, or comparable technologies.
  • Hands-on experience with virtualization platforms such as Xen, VMware, KVM, or comparable technologies.
  • Strong understanding of storage protocols, multipathing, volumes, filesystems, snapshots, replication, performance, capacity, and data-protection concepts.
  • Strong Linux administration and troubleshooting across services, filesystems, devices, networking, permissions, performance, and logs.
  • Experience with server hardware, hypervisors, clustering, virtual-machine lifecycle operations, and infrastructure management tooling.
  • Automation experience with Python, shell, Ansible, REST APIs, Git, CI/CD, or configuration-management systems.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Zabbix, or comparable platforms.
  • Experience with Jira or comparable ticketing systems, major-incident response, change control, postmortems, and problem management.
  • Ability to work independently during assigned offshore coverage and provide complete handoffs to United States-based IOQ and engineering teams.

ACADEMIC CREDENTIALS:

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field preferred; equivalent relevant experience considered.
  • Storage, virtualization, Linux, cloud, automation, or IT service-management certification is beneficial.

LOCATION: Austin, TX

This role is not eligible for visa sponsorship.

#LI-TL1

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against AMD's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on AMD's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    AMD's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.