Staff Production Engineer
Crusoe is an AI infrastructure company that designs, builds, and operates AI data centers and a cloud platform. It provides managed AI services, GPU compute, model fine-tuning and inference, and infrastructure operations for organizations building and deploying AI workloads.
Maintainer signals as of 8/14/2026
Funding history
Projects
About Crusoe
Crusoe, the AI factory company, provides Crusoe Cloud and Crusoe Intelligence Foundry for AI development and production. Its offerings include managed inference, serverless fine-tuning, high-performance NVIDIA and AMD compute, accelerated storage, RDMA networking, managed Kubernetes and Slurm, and operations tooling. The company also designs, builds, and operates modular AI data-center infrastructure using an energy-first approach, serving customers that need scalable training, inference, and AI platform infrastructure.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will maintain reliable production infrastructure by analyzing capacity, troubleshooting network and hardware issues, supporting data center builds, managing hardware replacements and RMAs, diagnosing optical network problems, improving operational processes, and collaborating on network architecture and automation.
Requirements
- 5+ years of professional production engineering experience
- 5+ years contributing to architecture and design of new and current systems
- Bachelor's degree in Computer Science or a related field, or 8+ years of relevant work experience
- Experience writing high-quality code in Python, Go, or a similar language
- Experience with Docker, Kubernetes, Ansible, Cloud Formation, and Terraform
- Experience with modern CI/CD practices and build systems
- Experience with logging, monitoring, and alerting systems
- Experience with Unix/Linux environments
- Experience with TCP/IP and network programming
- Experience with information security best practices
- Solid understanding of IP subnetting, Layer 2 and Layer 3 networking, VLANs, MAC addresses, port speeds, optics, and routing
- 2+ years of experience with BGP, MPLS, L3VPN, VPLS, Multicast, CoS, TCP, and IPv4/IPv6
- 2+ years of experience as a Network Engineer for a content or network provider
- Experience with data center network architectures including CLOS
- Experience with structured cabling in a data center environment
- Experience managing infrastructure and logistics vendors
Responsibilities
- Analyze network capacity and co-develop growth plans
- Contribute to network builds in data centers and Points of Presence
- Provide Tier 1 troubleshooting for switch provisioning and network-related server build failures
- Manage console DNS mapping and troubleshoot connectivity issues
- Verify automation scripts and conduct network audits
- Manage network hardware failures, switch swaps, appliance replacements, and RMA processes
- Troubleshoot Layer 1 optical network issues using OTDR tests
- Partner on structural cabling and fabric switch uplifts
- Lead operational excellence projects for recurring network issues
- Conduct network studies and recommend network improvements
- Evaluate and adopt new technologies and methodologies
Benefits
- Full social security coverage
- Contributions to provident, trade union, and pension funds
- Options for additional pensions
- Optional global life insurance
- Optional private health insurance
- Maternity leave
- Paternity leave
- Parental leave
- Sick leave
