Staff Technical Program Manager, Managed Intelligence
Crusoe is an AI infrastructure and cloud computing company. It provides GPU compute, AI model training and inference, developer tools, managed orchestration, data centers, and energy infrastructure for AI builders and enterprise customers.
Maintainer signals as of 8/23/2026
Funding history
Projects
About Crusoe, Inc
Crusoe designs, builds, and operates energy-first AI infrastructure, including high-performance data centers, modular AI factories, and Crusoe Cloud. Its cloud platform provides NVIDIA and AMD GPU compute, scalable storage, high-performance networking, managed Kubernetes and Slurm, observability, model training, fine-tuning, and inference services. Crusoe serves AI developers, startups, enterprises, and teams running production-scale AI workloads.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
Connect model engineering, IaaS, product, and data center operations to deliver a reliable and scalable managed inference platform. Own multi-quarter release planning, model onboarding, inference optimization, production readiness, capacity planning, risk identification, execution frameworks, dashboards, and executive communication.
Requirements
- 7+ years of experience as a Technical Program Manager
- Experience owning complex programs across engineering and product organizations
- Knowledge of LLM inference and model serving
- Familiarity with batching, quantization, latency, throughput, and production cost tradeoffs
- Experience with multi-tenant systems, isolation, quota management, and SLA enforcement
- Familiarity with fine-tuning and alignment workflows
- Ability to build execution models where processes do not yet exist
- Exceptional executive communication
- Daily use of AI tools for program execution
- Experience driving cross-functional alignment without direct authority
Responsibilities
- Own multi-quarter release planning and dependency governance
- Drive model version rollouts and inference optimization campaigns
- Manage SLA readiness for new GPU hardware
- Plan multi-tenant capacity
- Coordinate Model Engineering, IaaS, Cloud Foundations, Data Center Operations, and model providers
- Identify risks across model serving, reliability, capacity, and vendors
- Build execution frameworks and maintain real-time dashboards
- Plan model onboarding on new GPU generations
- Validate firmware, drivers, CUDA, ROCm, and inference commissioning criteria
- Drive alignment across engineering, product, and operations leadership
Benefits
- Competitive compensation and equity packages
- Restricted Stock Units
- Paid time off
- Paid holidays
- Leave of absence programs
- Comprehensive health insurance
- Dental insurance
- Vision insurance
- Employer contributions to HSA account
- Paid parental leave
- Paid life insurance
- Short-term disability
- Long-term disability
- Professional development
- Tuition reimbursement
- Mental health and wellness support
- Commuter benefits
- Cell phone stipend
- 401(k) retirement plan with company match up to 4% of salary
- Volunteer time off
- Global travel insurance and emergency assistance
- Daily meals allowance
- Additional location-specific perks and programs
