Head of Engineering GPU
European sovereign cloud and AI provider offering GPU compute, AI model APIs, cloud infrastructure, bare metal, managed data, containers, and quantum services.
About Scaleway
Scaleway SAS is a Paris-based subsidiary of iliad Group that operates a European cloud and AI platform. Its current offering includes GPU clusters and instances, serverless generative-model APIs, compute, storage, networking, managed databases, Kubernetes, serverless services, and Quantum as a Service.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will lead Support Engineering, HPC, and SRE organizations and manage two Engineering Managers. You will own GPU-cloud technical strategy and architecture, validate commercial commitments, oversee GPU-cluster deployment and readiness, improve automation and operations, and ensure infrastructure reliability, scalability, and performance.
Requirements
- 10+ years of infrastructure engineering experience
- Experience leading senior technical teams and managers
- Experience with large-scale infrastructure and compute clusters
- Knowledge of GPU or HPC environments
- Knowledge of NVIDIA or AMD technologies
- Knowledge of distributed infrastructure and cluster architecture
- Knowledge of reliability and production operations
- Experience with Kubernetes, Proxmox, Warewulf, Prometheus, and Grafana
- Experience with Lustre, DDN, and VAST Data
- Architecture trade-off analysis
- Stakeholder management
Responsibilities
- Lead Support Engineering, HPC, and SRE organizations
- Manage two Engineering Managers
- Own GPU Cloud technical strategy and architecture
- Validate technical and service dimensions of commercial proposals
- Oversee GPU cluster design, deployment, and operational readiness
- Improve cluster management, capacity management, automation, and operations
- Ensure infrastructure reliability, scalability, performance, and maintainability
- Provide technical leadership for AI and HPC infrastructure projects
- Align Engineering, GTM, Product, and Operations
- Develop the engineering organization through delegation and coaching
- Maintain engineering and operational standards
Benefits
- Up to 3 days of remote work per week
- Office outdoor spaces and bike parking
- Healthy meal service and breakfast
- Swile lunch card for regional-site employees
- Access to a gym
- Daycare places
- Discounted caring services
Hiring Process
HR discovery call; technical-skills interview; CTO technical interview; management interview; HR office interview.
