Search...

Principal Systems Software Engineer

Crusoe logo
Crusoe

Crusoe is an AI infrastructure company that designs, builds, and operates AI data centers and a cloud platform. It provides managed AI services, GPU compute, model fine-tuning and inference, and infrastructure operations for organizations building and deploying AI workloads.

Series C13 current maintainers10 active leads5 new active leads9 lead step-downs1 early lead departureTeam intelligence

Maintainer signals as of 8/12/2026

Distributed
About Crusoe

Crusoe, the AI factory company, provides Crusoe Cloud and Crusoe Intelligence Foundry for AI development and production. Its offerings include managed inference, serverless fine-tuning, high-performance NVIDIA and AMD compute, accelerated storage, RDMA networking, managed Kubernetes and Slurm, and operations tooling. The company also designs, builds, and operates modular AI data-center infrastructure using an energy-first approach, serving customers that need scalable training, inference, and AI platform infrastructure.

View jobs by Crusoe

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead the design of next-generation AI infrastructure spanning bare-metal, virtualized, and containerized compute. You will shape the I/O path, advise on hardware and software co-design, lead research and development workstreams, write technical strategy documents, debug kernel-level systems, optimize GPU cluster performance, and represent the work in open-source and industry forums.

Requirements

  • 12+ years designing and shipping core infrastructure at a major hyperscaler or specialized HPC cloud
  • Authoritative knowledge of the Linux kernel
  • Knowledge of KVM, QEMU, and Firecracker
  • Knowledge of RoCE v2 and InfiniBand
  • Experience designing software for NVIDIA or AMD GPUs and high-speed NICs
  • Experience leading cross-functional teams through high-ambiguity projects
  • Experience delivering production-ready mission-critical systems
  • Significant industry contributions such as patents, open-source contributions, or published research
  • Bachelor's or master's degree in computer science, computer engineering, or a related analytical field, or equivalent professional experience

Responsibilities

  • Architect bare-metal GPU infrastructure over InfiniBand and RDMA fabrics
  • Design thin virtualization layers using KVM or custom micro-VMs
  • Build container infrastructure using Kubernetes or Slurm
  • Lead the architecture of the internal cloud fabric
  • Drive the roadmap for SR-IOV, RDMA, and virtualized GPU scheduling
  • Lead research and development workstreams
  • Draft white papers and RFCs
  • Debug race conditions in the I/O path
  • Optimize kernel-level memory pinning for GPU clusters
  • Represent Crusoe in open-source communities and industry forums

Benefits

  • Competitive compensation
  • Restricted Stock Units
  • Paid time off
  • Paid holidays
  • Comprehensive health insurance
  • Dental insurance
  • Vision insurance
  • Employer HSA contributions
  • Paid parental leave
  • Paid life insurance
  • Short-term disability insurance
  • Long-term disability insurance
  • Professional development
  • Tuition reimbursement
  • Mental health and wellness support
  • Commuter benefits
  • Cell phone stipend
  • 401(k) retirement plan with company match up to 4% of salary
  • Volunteer time off