Senior hands-on engineer who leads break-fix and critical-incident response across customer data centers. Fixes the hard problems, runs the bridge during outages, and makes sure failures don't repeat.
Key Responsibilities
- Lead break-fix and critical-incident response in office—triage, troubleshoot, and resolve.
- Act as a technical lead on Sev-1/major-incident bridges; coordinate field engineers, TAC, OEM vendors, and the customer until it's closed.
- Work hands-on across servers, storage, network hardware, structured cabling (copper/fiber/MPO), power/PDU, and rack & stack.
- Run RCA on major incidents and close the corrective actions so the same fault doesn't recur.
- Support hardware replacements, upgrades, and migrations.
- Document fixes and keep runbooks/SOPs current so the team resolves faster next time.
- Mentor field engineers on tougher diagnostics.
Education & Experience
- 5–8 years of hands-on data center/infrastructure field experience; 3+ years in break-fix (enterprise or hyperscale).
- Strong troubleshooting across server, network, cabling, and power — with good judgment on when to escalate.
- Able to own a Sev-1 bridge under pressure with the customer on the line.
- Working knowledge of ServiceNow / Jira ticketing.
- A degree is helpful, not required—field experience counts.