Data Center Site Manager / Supervisor
Bitdeer is a technology company providing Bitcoin mining solutions.
Funding history
Investors
Projects
About Bitdeer
Bitdeer provides full-spectrum Bitcoin mining and high-performance computing solutions, including SEALMINER mining equipment, Minerbase cooling containers, cloud mining, co-mining, mining management applications, mining rights marketplaces, and large-scale data center operations. The company also offers AI cloud infrastructure with GPU computing, model training and deployment capabilities, and turnkey AI data center solutions for enterprise customers and developers. Bitdeer is headquartered in Singapore and operates globally.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will lead daily data center operations and manage Operations Engineers. You will plan staffing and shifts, oversee AI/HPC infrastructure, handle escalations, improve procedures, review maintenance and incident work, and coordinate deployments. You will participate in on-call support and ensure operational, safety, and security compliance.
Requirements
- Bachelor's degree or above in a related technical discipline
- At least 5 years of data center operations, IT infrastructure, or HPC/AI infrastructure management experience
- At least 2 years of team leadership or people-management experience
- Knowledge of data center operations and AI/HPC infrastructure
- Familiarity with NVIDIA GPU architecture, NVLink, NVSwitch, and AI cluster deployment
- Experience with hardware troubleshooting, firmware management, and hardware lifecycle management
- Understanding of structured cabling systems
- Linux administration and troubleshooting skills
- Experience with shift scheduling, workforce planning, incident management, performance management, operational governance, and vendor coordination
Responsibilities
- Lead daily data center site operations
- Manage Operations Engineers through staffing, scheduling, assignments, performance management, coaching, and development
- Ensure 24x7 operational coverage
- Act as the primary escalation point and coordinate incident resolution and root-cause analysis
- Oversee operation, maintenance, and troubleshooting of AI/HPC infrastructure
- Establish and improve SOPs, EOPs, and preventive-maintenance programs
- Monitor site health, KPIs, incident trends, and infrastructure performance
- Coordinate installations, commissioning, expansions, and lifecycle management
- Review maintenance activities, change requests, incident reports, and shift handovers
- Ensure compliance with operational, safety, and security requirements
- Coordinate with engineering, network, facilities, and vendor teams
- Participate in on-call rotation and provide hands-on support
- Drive operational excellence and continuous improvement
