Skip to content

Data Replication Specialist

IBM

Herndon, US

Applying for this one?

We write the CV against this exact posting — its wording, its requirements — not a template with your name in it.

Get my CV for this job

$25, one-time. No subscription.

This position is for a Data Replication Specialist supporting the Army Edge Computing Capability (AECC) project that ALTESS is fielding for the US Army. The AECC solution is a containerized, Kubernetes-based, multitenant hosting environment for hosting Army enterprise and tactical applications. AECC is utilizing Kubernetes and potentially Red Hat OpenShift to implement a cloud-native, software-defined infrastructure across multiple global sites. ALTESS is utilizing enterprise backup software solutions (e.g., Kasten K10, Velero, TrilioVault, CloudCasa, Cohesity DataProtect, or similar) to provide enterprise backup capabilities and replication of backup data to remote sites to support mission needs. ALTESS provides value-added common and managed services built on top of the Kubernetes foundation that hosted Army applications will require. ALTESS is a managed service provider (MSP) and hosting services provider for Army applications. ALTESS is a Product Lead office under Capability Program Executive (CPE) Enterprise Software and Services (CPE ES2).

Position Duties:

• Develop an enterprise-level backup and data replication solution utilizing Kubernetes-native backup tools for all sites within the AECC environment.

• Automate backup, replication, and recovery processes for containerized applications, Kubernetes clusters, and virtual machines hosted within the AECC environment.

• Configure and manage Kubernetes-native backup tools to ensure seamless integration with Kubernetes clusters, including support for persistent volumes, CSI (Container Storage Interface), and application-aware backups.

• As new applications are hosted on the AECC solution, utilize Infrastructure as Code (IaC) for the configuration and administration of backups within the hosted customer enclaves.

• Provide daily administration of backup tools, including patching, upgrading, monitoring, and optimizing backup schedules.

• Monitor utilization of backup resources and notify the AECC Lead when additional resources will be required.

• Harden backup tools and processes per commercial best practices and required government cybersecurity controls.

• Troubleshoot and perform root cause analysis of backup-related issues across Kubernetes clusters, containers, and virtual machines.

• Develop and maintain backup and recovery process documentation (system diagrams, network topology, software configurations, device configurations, standard operating procedures).

• Provide on-call support for triage and resolution of after-hours production incidents.

• Make recommendations for improvements to backup and recovery capabilities and services provided to hosted applications.

Required Skills:

• Senior-level experience with enterprise backup, replication, and recovery solutions in a Kubernetes based environment

• Experience with backup, replication, and recovery processes in a multi-site Kubernetes deployment.

• Strong automation and Infrastructure as Code (IaC) skills, including tools such as Terraform, Ansible, Vault, etc.

Required Certifications:

• Security+ or equivalent DoD 8570.01-M IA Tech Level II certification.

• Must have (or obtain within 6 months of hire) a computing environment certification as defined in DoD 8570.01-M, such as an enterprise backup, replication, or recovery vendor's certification.

Clearance Required:

• DoD Secret

Position Location:

• (Hybrid telework/onsite as needed in Radford, VA )

Education:

• Bachelor's degree or higher in IT related field

Desired Skills:

• Working knowledge of DoD Security Technical Implementation Guides (STIG) and the Information Assurance Vulnerability Management (IAVM) process, and/or industry hardening best practices and processes.

• Knowledge and experience with containerization and container orchestration tools.

• Proficiency in scripting languages (Python, PowerShell, BASH, etc.) for automation tasks.

• Strong troubleshooting skills across the entire technology stack – network, storage, server, containers, and applications.

Seen 10 hours ago · IBM postings close after a median of 2 days.

Original posting on IBM's site ↗

Posting text belongs to the employer. Removal requests: contact us.

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

One job at a time

One posting. One CV. $25.

Pick the job you actually want and we write for it.

Get my CV for this job