Member of Technical Staff Datacenter Networking
Prime Intellect is an active AI infrastructure company building an open stack for training, deploying, evaluating, and continuously improving agentic models.
Maintainer signals as of 9/2/2026
Funding history
Projects
About Prime Intellect
Prime Intellect, Inc. operates a full-stack AI platform combining hosted reinforcement-learning training, evaluations, inference, secure sandboxes, GPU compute, and open-source research tooling. Its current positioning is the Open Superintelligence Stack, serving AI companies, researchers, and developers.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will design and operate networks for GPU training, inference, storage, and management workloads. You will automate network operations, diagnose performance and reliability issues, benchmark fabrics, and improve monitoring and incident response.
Requirements
- 3+ years of production datacenter networking experience
- Understanding of Ethernet, TCP IP, routing, switching, and redundant network design
- Experience with InfiniBand or RoCE GPU networking
- Experience troubleshooting Linux hosts, NICs, switches, and physical links
- Experience automating network operations with Python, Ansible, or comparable tools
- Knowledge of leaf-spine architectures, BGP, ECMP, VLANs, RDMA, Linux networking, telemetry, and alerting
Responsibilities
- Design and deploy scalable datacenter network topologies
- Configure and operate Ethernet RoCE and InfiniBand fabrics
- Automate network provisioning, configuration validation, upgrades, and rollbacks
- Diagnose packet loss, congestion, link failures, and collective communication performance
- Benchmark end-to-end network performance
- Build monitoring for port health, errors, utilization, congestion, and fabric topology
- Partner on cabling, optics, deployment readiness, and failure resolution
Benefits
- Equity incentives
