Network Operations Center Manager
Nscale is a London-based, full-stack AI cloud and infrastructure company that provides GPU compute, managed AI services, orchestration software, data centers, and power infrastructure for AI training, fine-tuning, and inference.
Funding history
About Nscale
Nscale builds and operates vertically integrated AI infrastructure spanning software, GPU compute, networking, storage, purpose-built data centers, and power. Its active cloud platform offers self-service inference endpoints, fine-tuning, managed Kubernetes and Slurm, virtual machines, and GPU clusters.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will lead 24/7 network operations at the Glomfjord data centre. You will manage staff, monitor network and facility infrastructure, lead incident response, oversee BMS operations, report performance, maintain procedures, coordinate maintenance, improve automation, and support compliance and continuity planning.
Requirements
- 4+ years managing or leading a team in a NOC, TAC, or critical data centre environment
- 6+ years of hands-on experience in network engineering, systems administration, or infrastructure operations
- Bachelor's degree in Computer Science, Network Engineering, Information Technology, or equivalent practical experience
- Knowledge of BGP, OSPF, switching, firewalls, VPNs, and load balancing
- Proficiency with monitoring, SIEM, and ticketing tools
- English communication skills
- Willingness to be site-based in Norway
Responsibilities
- Lead, develop, and manage a 24/7 Network Operations Centre team
- Create shift schedules for continuous operational coverage
- Oversee monitoring of network infrastructure, servers, customer environments, BMS, and critical infrastructure
- Lead critical incident response and root-cause analysis
- Oversee BMS monitoring for electrical, cooling, environmental, fire, and access-control systems
- Produce environmental and operational performance reports
- Monitor and report SLAs, KPIs, uptime, incident response, and ticket resolution
- Develop and maintain operating procedures, runbooks, monitoring standards, and escalation workflows
- Coordinate planned maintenance, upgrades, and emergency works
- Ensure comprehensive shift handovers
- Improve monitoring, automation, workflows, and operational tooling
- Support business continuity, disaster recovery, audits, and compliance reviews
