
Senior Infrastructure Engineer
Harrison Clarke · United States
Remote
About the job
AI Infrastructure Engineer
We're partnered with a fast-growing company building the infrastructure layer for the global GPU economy. They aggregate GPU capacity across cloud providers and data centers into a unified platform. This is a lean team that moves fast. They want the same from you.
The Role
You'll own the infrastructure layer - Kubernetes clusters, bare metal provisioning, and the tooling that keeps heterogeneous GPU environments running reliably at scale. This is not an ML role. You won't be training models. You'll be making sure the infrastructure underneath them doesn't fall over.
What You'll Work On
Managing and scaling Kubernetes clusters across diverse hardware, geographies, and providers
Bare metal provisioning and low-level systems work - IPMI, BMC/Redfish, physical server management
Building automation tooling to standardize deployments across heterogeneous environments
Deep debugging across inconsistent infrastructure setups
On-call participation - this is an infrastructure company, incidents happen, ownership is expected
What We're Looking For
Strong Kubernetes experience in production, multi-tenant or multi-provider environments
Genuine bare metal experience - not abstracted cloud layers, but physical server provisioning and management (IPMI/BMC/Redfish familiarity is a strong signal)
Systems-level debugging instincts and solid root-cause analysis skills
Comfortable operating with less abstraction and more ambiguity than a typical platform role
GPU provisioning or ML infrastructure experience is a nice-to-have, not a requirement - we care about infra depth, not model development!
$200K/yr - $250K/yr · 12 benefitsApply now