Compute Engineer Deployment

AI infrastructure company that builds and operates large-scale compute and data-center infrastructure for frontier AI workloads.

New York City, United States
About Fluidstack

Fluidstack deploys AI compute infrastructure, including custom data centers and large-scale compute capacity, for AI labs, governments, and enterprises.

View jobs by Fluidstack

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will own compute turn-up from facility availability through ready-for-service. You will qualify racks, configure firmware and out-of-band management, validate nodes and clusters, automate provisioning workflows, triage hardware failures, and support deployment and incident response during turn-up windows.

Requirements

  • Experience bringing server or GPU fleets of hundreds of nodes or more to production
  • Linux, BMC, IPMI, and Redfish expertise
  • Experience automating hardware workflows in Python or Go
  • Hands-on experience racking, cabling, and swapping data center components
  • Ability to triage hardware, firmware, and software failures
  • Ability to travel for turn-up windows

Responsibilities

  • Own compute turn-up from facility availability to ready-for-service
  • Qualify racks by establishing firmware baselines, configuring BMC and BIOS, running burn-in, and validating nodes and clusters
  • Drive qualification through Kubernetes-based provisioning and shared services
  • Triage hardware failures and manage RMA and vendor escalation
  • Run remote turn-up and perform on-site deployment visits
  • Partner with deployment and operations functions during turn-up windows
  • Support incident response on newly live capacity

Benefits

  • Equity
Compute Engineer Deployment at Fluidstack | JobStash