Who are we?
Equinix is the world’s digital infrastructure company®, shortening the path to connectivity to enable the innovations that enrich our work, life and planet.
A place where bold ideas are welcomed, human connection is valued, and everyone has the opportunity to shape their future.
Help us challenge assumptions, uncover bias, and remove barriers—because progress starts with fresh ideas. You’ll find belonging, purpose, and a team that welcomes you—because when you feel valued, you’re empowered to do your best work.
Job Summary
We are looking for a Staff Engineer to join the Reliability Engineering team that operates a bare-metal Kubernetes platform across multiple metro locations. The platform's architecture is defined and the build is underway you will provision and operate it day to day: bringing up nodes and clusters, running upgrades, keeping the fleet consistent through GitOps, and owning your shift of the on-call rotation.
This is a deliberately broad role. We are not looking for someone who works one layer and hands everything below it across a boundary.
You will work alongside the teams that own the physical hardware, the underlay network, and the data plane software, and you need working knowledge of all three enough to isolate where a problem lives before escalating it.
You will join the Digital Interconnection Engineering organization. The team owns the full application stack for interconnection and the Kubernetes platform for the software-defined data plane across multiple metros. We operate with an automation-first, GitOps-driven culture and a strong commitment to operational rigor, documented runbooks, and blameless postmortems.
Responsibilities
OS Provisioning & Bare-Metal Infrastructure
Provision bare-metal servers across multiple metros using the team's PXE/iPXE imaging and cloud-init pipeline, and harden the OS prior to cluster bootstrap
Extend and maintain that provisioning automation as current hardware, metros and OS versions are introduced,