Data Engineer

AI infrastructure company that builds and operates large-scale compute and data-center infrastructure for frontier AI workloads.

New York City, United States
About Fluidstack

Fluidstack deploys AI compute infrastructure, including custom data centers and large-scale compute capacity, for AI labs, governments, and enterprises.

View jobs by Fluidstack

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build production data pipelines that consolidate company systems into a queryable layer. You will own data models for the live knowledge graph, deliver reliable datasets and services, and transform vendor and field data into structured, trustworthy inputs.

Requirements

  • Experience building and operating production data pipelines
  • Experience modeling messy real-world domains into durable schemas
  • Knowledge of data quality testing, monitoring, and lineage
  • Experience extracting structure from unstructured sources
  • Experience with AI tools and modern data stacks

Responsibilities

  • Build pipelines that consolidate company systems into a queryable data layer
  • Own data models for sites, equipment, schedules, and people
  • Ship datasets and services with SLAs for internal tools, dashboards, and ML models
  • Transform unstructured vendor and field data into reliable structured inputs

Benefits

  • Equity