Search...

Staff Network Production Engineer Network Ops

Crusoe logo
Crusoe

Crusoe is an AI infrastructure company that designs, builds, and operates AI data centers and a cloud platform. It provides managed AI services, GPU compute, model fine-tuning and inference, and infrastructure operations for organizations building and deploying AI workloads.

Maintainer signals as of 8/14/2026

Distributed
About Crusoe

Crusoe, the AI factory company, provides Crusoe Cloud and Crusoe Intelligence Foundry for AI development and production. Its offerings include managed inference, serverless fine-tuning, high-performance NVIDIA and AMD compute, accelerated storage, RDMA networking, managed Kubernetes and Slurm, and operations tooling. The company also designs, builds, and operates modular AI data-center infrastructure using an energy-first approach, serving customers that need scalable training, inference, and AI platform infrastructure.

View jobs by Crusoe

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will operate Crusoe’s global edge, backbone, and data center network supporting GPU-based HPC clusters. You will monitor performance, troubleshoot incidents, execute network changes, improve reliability and observability, mentor network engineers, manage vendors, analyze network health, and participate in 24/7 on-call support.

Requirements

  • 10+ years of related experience operating at scale in a production environment
  • Knowledge of TCP/IP, QoS, BGP, OSPF/IS-IS, EVPN, VXLAN, RSVP-TE, and LDP
  • Understanding of SNMP, IPFIX, sFlow/netflow, and telemetry
  • Familiarity with data center, backbone, and edge network architecture
  • Experience with Python or similar languages for automation and tooling
  • Experience with Mellanox, Cisco, Arista, Juniper, and other network devices
  • Familiarity with Broadcom and Barefoot switch and router chipsets
  • Knowledge of public cloud connectivity options
  • Understanding of IPv6 and IPv4-IPv6 coexistence technologies
  • Experience with network observability, flow analytics, and dashboarding tools
  • Bachelor’s degree in a relevant field or equivalent experience

Responsibilities

  • Monitor network performance
  • Perform advanced troubleshooting and root-cause analysis
  • Guide post-mortem reviews and improvements
  • Execute network changes across data centers, backbone, and edge infrastructure
  • Manage and optimize the global network
  • Collaborate with Network Engineering and cross-functional teams
  • Develop monitoring, alerting, and network availability systems
  • Mentor network engineers
  • Establish incident response and operational readiness practices
  • Manage network vendors and contracts
  • Deliver network health metrics, event statistics, and performance analysis
  • Participate in 24/7 network on-call support

Benefits

  • Pension contributions
  • Private health insurance
  • Dental insurance
  • Income protection
  • Life assurance