Senior Hardware Systems Engineer
Crusoe is an AI infrastructure company that designs, builds, and operates AI data centers and a cloud platform. It provides managed AI services, GPU compute, model fine-tuning and inference, and infrastructure operations for organizations building and deploying AI workloads.
Maintainer signals as of 8/12/2026
Funding history
Projects
About Crusoe
Crusoe, the AI factory company, provides Crusoe Cloud and Crusoe Intelligence Foundry for AI development and production. Its offerings include managed inference, serverless fine-tuning, high-performance NVIDIA and AMD compute, accelerated storage, RDMA networking, managed Kubernetes and Slurm, and operations tooling. The company also designs, builds, and operates modular AI data-center infrastructure using an energy-first approach, serving customers that need scalable training, inference, and AI platform infrastructure.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will support the full hardware lifecycle from prototype bring-up through production. You will build testing automation, validate GPU and CPU platforms, debug PCIe, InfiniBand, and NVMe systems, investigate failures, support integration and production readiness, and resolve system-level issues across engineering and manufacturing functions.
Requirements
- 4-8 years of experience in hardware development, validation, sustaining engineering, or production engineering
- Hands-on expertise in PCIe, InfiniBand, and NVMe/storage debugging and development
- Proficiency in hardware bring-up, board-level debugging, and system-level validation
- Ability to implement hardware testing automation using Python, Shell, or similar languages
- Background in digital and analog design, server architecture, and high-performance computing hardware
- Experience working across thermal, mechanical, firmware, and software functions
- Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, or equivalent experience
Responsibilities
- Drive the hardware development and sustaining lifecycle
- Develop and maintain hardware testing and diagnostics automation
- Debug PCIe link training, topology, and performance issues
- Debug InfiniBand fabric, throughput, and connectivity issues
- Investigate NVMe and storage bottlenecks, firmware interactions, and failures
- Validate and characterize GPU, CPU, and high-performance computing platforms
- Support end-to-end integration and solution testing
- Resolve system-level issues with multidisciplinary teams
- Drive prototyping, qualification, and high-volume manufacturing readiness
- Provide data-driven insights for hardware roadmaps and reliability strategy
Benefits
- Restricted Stock Units
