Senior GPU Systems & Fabric Engineer

Bitdeer is a technology company providing Bitcoin mining solutions.

0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/25/2026

Singapore, SG
About Bitdeer

Bitdeer provides full-spectrum Bitcoin mining and high-performance computing solutions, including SEALMINER mining equipment, Minerbase cooling containers, cloud mining, co-mining, mining management applications, mining rights marketplaces, and large-scale data center operations. The company also offers AI cloud infrastructure with GPU computing, model training and deployment capabilities, and turnkey AI data center solutions for enterprise customers and developers. Bitdeer is headquartered in Singapore and operates globally.

View jobs by Bitdeer

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

Architect GPU device plugin and Kubernetes Operator integrations; optimize RDMA, SR-IOV, RoCEv2, and InfiniBand networking; build automated GPU and NIC remediation pipelines; manage MIG and vGPU slicing; tune kernels, drivers, CUDA, and NCCL; support topology-aware placement and data movement; define bare-metal provisioning and hardening standards; investigate performance issues; and mentor team members.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or a related field
  • 5+ years of systems engineering experience
  • Strong proficiency in Linux kernel internals, C, or Go
  • Hands-on experience with NVIDIA H100/A100 GPU architectures
  • Experience with CUDA runtimes and distributed networking
  • Understanding of containerized environments and Kubernetes device plugin architecture
  • Experience operating, debugging, and scaling bare-metal systems
  • Familiarity with Terraform, Ansible, and CI/CD pipelines
  • Strong problem-solving skills
  • Excellent communication skills

Responsibilities

  • Architect and maintain NVIDIA and AMD GPU device plugin and Kubernetes Operator integrations
  • Configure and optimize RDMA, SR-IOV, RoCEv2, and InfiniBand networking
  • Build automated hardware remediation pipelines using DCGM telemetry
  • Manage MIG and vGPU slicing technologies
  • Tune kernel parameters, device drivers, CUDA, and NCCL
  • Collaborate on topology-aware placement and data movement
  • Define standards for bare-metal provisioning, BIOS and firmware updates, and OS hardening
  • Lead investigations into hardware, fabric, and software performance issues
  • Mentor team members and drive documentation standards

Benefits

  • A culture that values authenticity and diversity of thoughts and backgrounds
  • An inclusive and respectable environment with open workspaces and exciting start-up spirit
  • Fast-growing company with the chance to network with industrial pioneers and enthusiasts
  • Ability to contribute directly and make an impact on the future of the digital asset industry
  • Involvement in new projects and developing processes and systems
  • Personal accountability, autonomy, fast growth, and learning opportunities
  • Attractive welfare benefits and developmental opportunities such as training and mentoring