Search...

Senior Engineering Manager, SDN Control Plane

Crusoe logo
Crusoe

Crusoe is an AI infrastructure company that designs, builds, and operates AI data centers and a cloud platform. It provides managed AI services, GPU compute, model fine-tuning and inference, and infrastructure operations for organizations building and deploying AI workloads.

Series C13 current maintainers10 active leads5 new active leads9 lead step-downs1 early lead departureTeam intelligence

Maintainer signals as of 8/12/2026

Distributed
About Crusoe

Crusoe, the AI factory company, provides Crusoe Cloud and Crusoe Intelligence Foundry for AI development and production. Its offerings include managed inference, serverless fine-tuning, high-performance NVIDIA and AMD compute, accelerated storage, RDMA networking, managed Kubernetes and Slurm, and operations tooling. The company also designs, builds, and operates modular AI data-center infrastructure using an energy-first approach, serving customers that need scalable training, inference, and AI platform infrastructure.

View jobs by Crusoe

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead the team building the scalable SDN control plane for VPC networking across large GPU fleets. You will own the roadmap, architecture, implementation, and production operation of distributed systems for network virtualization, guide reliability and scale work, mentor senior and staff engineers, and collaborate with data-plane and cloud product teams.

Requirements

  • 10+ years in distributed systems or cloud networking engineering
  • 5-7+ years managing senior and staff-level talent
  • Deep knowledge of SDN and network virtualization
  • Knowledge of VXLAN, Geneve, VPC constructs, BGP, and EVPN
  • Experience with OVN, OVS, or equivalent control-plane architectures
  • Hands-on experience building large-scale control planes
  • Experience with state reconciliation, consensus, consistency, API design, and fleet-wide configuration propagation
  • Experience with Go, Kubernetes-style controllers, or similar technologies
  • Understanding of SLOs, convergence-time benchmarking, scale benchmarking, graceful degradation, and blast-radius containment

Responsibilities

  • Define the VPC control-plane roadmap
  • Lead the evolution of network virtualization systems
  • Oversee distributed control-plane service design
  • Drive scalability beyond the current fleet size
  • Lead reliability engineering and API-latency benchmarking
  • Prevent regressions and respond to incidents
  • Mentor and grow senior and staff distributed-systems engineers
  • Set technical standards and accountability
  • Partner with data-plane and cloud product teams
  • Deliver low-latency, highly available networking for multi-tenant GPU clusters

Benefits

  • Competitive compensation and equity packages
  • Restricted Stock Units
  • Paid time off
  • Paid holidays
  • Leave of absence programs
  • Comprehensive health insurance
  • Dental insurance
  • Vision insurance
  • Employer HSA contributions
  • Paid parental leave
  • Paid life insurance
  • Short-term disability insurance
  • Long-term disability insurance
  • Professional development
  • Tuition reimbursement
  • Mental health and wellness support
  • Commuter benefits
  • Cell phone stipend
  • 401(k) retirement plan with company match up to 4% of salary
  • Volunteer time off
  • Global travel insurance and emergency assistance
  • Daily meals allowance
  • Location-specific perks and programs