DevOps GitLab-based Platform CI/CD 30*3 Pipelines
Bitdeer Technologies Group is a technology company providing Bitcoin mining solutions, mining hardware, data-center infrastructure, and AI cloud services. It serves individual, institutional, and enterprise customers globally.
Funding history
Investors
Projects
About Bitdeer Technologies Group
Bitdeer provides vertically integrated Bitcoin mining and high-performance computing services. Its operations include mining equipment procurement and manufacturing, datacenter design and construction, equipment management, daily mining operations, cloud mining, and mining-related services. The company also offers AI cloud infrastructure and high-performance computing powered by NVIDIA GPUs for AI and machine-learning workloads, serving customers across global markets.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will build and operate the CI/CD, infrastructure-as-code, MLOps, and internal developer platform foundations used to deploy AI products. You will manage Kubernetes and GPU infrastructure, improve availability and observability, enforce security controls, and lead incident resolution and preventative automation.
Requirements
- Bachelor’s degree or higher in Computer Science Engineering or a related technical field
- 5+ years of experience in DevOps Site Reliability Engineering or cloud infrastructure
- Expert knowledge of Linux and networking principles
- Deep Docker and Kubernetes expertise
- Experience with AWS GCP Azure Alibaba Cloud or hybrid cloud platforms
- Strong programming or scripting skills in Go Python Shell or another major language
- Understanding of CI/CD infrastructure as code observability and SRE principles
- Experience with MLOps model serving GPU clusters or AI infrastructure is preferred
- Experience with distributed systems platform engineering security or compliance is preferred
Responsibilities
- Design and maintain CI/CD pipelines for applications and machine learning models
- Build and scale Kubernetes Docker and GPU infrastructure
- Implement high availability disaster recovery and self-healing systems
- Automate infrastructure provisioning with Terraform Ansible and Helm
- Architect monitoring logging and alerting systems
- Build internal developer platform golden paths
- Establish security governance and compliance controls
- Lead incident troubleshooting root-cause analysis and remediation
