DATA CENTER OPERATIONS — DAY/FLEXIBLE
On-site in Pittsburgh, Pennsylvania | Monday–Friday, 9:00 a.m.–5:00 p.m. Eastern
This position normally works Monday through Friday from 9:00 a.m. to 5:00 p.m. Eastern. During another team member’s planned absence or a scheduled operational event, the employee may temporarily work alternate hours, including evening or weekend coverage.
Schedule changes will ordinarily be communicated in advance. When alternate coverage is required, the employee’s regular workdays or hours will be adjusted accordingly. This is not a permanently rotating-shift position.
WHO WE ARE
Teraswitch is on a mission to provide the highest performance, lowest latency bare metal servers in the world. With 20 data center locations around the world, Teraswitch has served thousands of customers across 185 countries.
Founded by Brendan Mannella and headquartered in Pittsburgh, Pennsylvania, Teraswitch is a privately held infrastructure company operating a growing global bare metal platform.
WHAT TO EXPECT
As a member of Data Center Operations, you will own infrastructure incidents, customer requests, maintenance activities, deployments, and operational projects from intake through verified completion.
This is primarily a remote infrastructure operations role—not a primarily hands-on data-center technician position. Most work is performed through monitoring systems, tickets, server-management interfaces, technical documentation, and coordination with customers, internal engineering teams, colocation facilities, vendors, contractors, and remote-hands providers.
Approximately 10% of the role may involve physical work including server assembly, component replacement, cabling, rack-and-stack work, and equipment maintenance.
This role requires someone who can investigate problems independently, communicate clearly with technical and nontechnical stakeholders, create precise technical procedures, manage multiple concurrent priorities, and remain accountable until work is completed and validated.
Responsibilities include:
- Own infrastructure incidents, customer requests, maintenance activities, deployments, and operational projects from intake through verified completion.
- Monitor infrastructure alerts, support requests, maintenance activities, and active operational issues.
- Investigate reported server, network, power, cooling, connectivity, deployment, and monitoring problems.
- Gather and interpret information from logs, monitoring platforms, tickets, management interfaces, internal systems, customers, and on-site technicians.
- Perform incident triage, take safe and documented corrective actions, and escalate service-affecting issues when appropriate.
- Communicate directly and professionally with customers during incidents, technical investigations, maintenance events, and service requests.
- Write precise Methods of Procedure, implementation plans, validation steps, rollback instructions, runbooks, and remote-hands directions.
- Translate technical requirements into clear, executable instructions for colocation technicians, contractors, vendors, and internal teams.
- Coordinate and direct remote-hands technicians at Teraswitch data center locations around the world.
- Remain accountable for remote work by confirming the correct equipment, reviewing evidence, tracking progress, validating results, and ensuring documentation is complete.
- Coordinate scheduled installations, migrations, maintenance events, hardware replacements, shipments, and other operational projects.
- Work with facility personnel to address power, cooling, cross-connect, cabling, access, and other facility-provided services.
- Coordinate vendor warranty technicians and ensure hardware issues are resolved promptly and correctly.
- Maintain clear ticket updates, incident timelines, project notes, escalation records, and actionable shift handoffs.
- Collaborate with Software, Network, Systems, Facilities, Logistics, and other internal teams.
- Maintain accurate records of infrastructure, hardware changes, completed work, and physical inventory.
- Identify recurring problems and improve documentation, procedures, tooling, and operational workflows.
- Execute documented server-management, recovery, deployment, and maintenance procedures.
- Perform occasional hands-on server assembly, rack-and-stack work, cabling, diagnostics, upgrades, and component replacement.
- Provide temporary coverage on alternate Operations schedules during planned team absences or scheduled operational events.
- Follow established security, change-management, escalation, validation, and access-control procedures.
WHAT OUR NEW TEAM MEMBER WILL NEED
- Three or more years of experience in data center operations, infrastructure support, systems support, technical operations, or a similar role.
- Demonstrated ability to write detailed technical procedures that another technician can execute without additional interpretation.
- Experience troubleshooting infrastructure remotely using logs, monitoring platforms, tickets, command-line tools, server-management interfaces, or information supplied by on-site personnel.
- Experience independently owning incidents, service requests, maintenance activities, or technical projects through completion.
- Strong written communication and the ability to provide clear updates to customers, vendors, contractors, and internal technical teams.
- Ability to prioritize multiple concurrent customer, infrastructure, and operational issues.
- Sound judgment regarding service impact, change control, escalation, rollback, and validation.
- Working knowledge of server hardware, network equipment, structured cabling, and data center facility services.
- Ability to distinguish between gathering information, taking corrective action, escalating an issue, and transferring ownership.
- Ability to follow technical procedures precisely and verify that completed work produced the intended result.
- Strong troubleshooting, organization, and time-management skills.
- Ability to safely lift, move, rack, and work around data center equipment when physical work is required.
- Ability to temporarily adjust scheduled workdays and hours, with advance notice when practicable, to provide coverage during planned team absences or scheduled operational events.
HELPFUL EXPERIENCE
- Writing MOPs, SOPs, runbooks, implementation plans, rollback procedures, and escalation documentation.
- Coordinating remotehands providers, contractors, vendors, and colocation facility personnel.
- Linux administration and troubleshooting.
- Server management through IPMI, Redfish, iDRAC, iLO, or other BMC platforms.
- Infrastructure monitoring and alert-management platforms.
- Ticketing, documentation, inventory, and infrastructure-management systems.
- Basic networking concepts, including IP addressing, VLANs, interface bonding, and physical connectivity.
- Server diagnostics, component replacement, firmware management, and hardware compatibility.
- Hardware inventory and spare-parts management.
- Experience supporting bare metal, hosting, cloud, network, or large-scale distributed infrastructure.
- Familiarity with APIs or API endpoints used to retrieve and validate operational data.
COMPENSATION AND BENEFITS
Along with competitive pay, full-time Teraswitch employees are eligible for the following benefits beginning on their first day of employment:
- Health, dental, and vision insurance.
- 401(k) with company profit sharing.
- Flexible paid time off.
- Company holiday benefits.
SUCCESS IN THIS ROLE
Success means operational issues receive prompt attention, technical evidence is gathered before potentially disruptive action is taken, customers and internal teams receive clear updates, and remote technicians receive instructions they can execute without ambiguity.
Work is not considered complete merely because it was escalated or assigned to another party. Success means the employee retains ownership, coordinates the required people, tracks progress, validates the result, updates the relevant documentation, and provides a complete and actionable handoff when work continues into another shift.
The employee also provides dependable temporary coverage when another Operations team member has a planned absence, while maintaining the same standards of ownership, communication, documentation, and judgment.
Seen 4 hours ago.
Original posting on teraswitch's site ↗
Posting text belongs to the employer. Removal requests: contact us.
Nearby
Live postings like this one
Same employer first, then the same role elsewhere.
- 4h ago
- 4h ago
- 4h ago
- Hybrid4h ago
- 4h ago
One job at a time
One posting. One CV. $25.
Pick the job you actually want and we write for it.