Site Reliability Engineer - SRE
European sovereign cloud and AI provider offering GPU compute, AI model APIs, cloud infrastructure, bare metal, managed data, containers, and quantum services.
Maintainer signals as of 9/25/2026
About Scaleway
Scaleway SAS is a Paris-based subsidiary of iliad Group that operates a European cloud and AI platform. Its current offering includes GPU clusters and instances, serverless generative-model APIs, compute, storage, networking, managed databases, Kubernetes, serverless services, and Quantum as a Service.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will automate monitoring, diagnosis, and incident remediation for production systems. You will troubleshoot high-impact issues, take part in on-call rotations, maintain observability and infrastructure tooling, document procedures, and collaborate with development teams on resilient infrastructure.
Requirements
- Experience with Go, Python, or Rust
- Strong Bash or Python scripting skills
- Hands-on Linux experience with Ubuntu or Debian
- Knowledge of TCP/IP, DNS, BGP, load balancing, and IPv6
- Experience with cloud infrastructure, bare metal, VMs, containers, and orchestrators
- Familiarity with Prometheus, Grafana, and Elastic
- Infrastructure-as-Code experience with Ansible, Salt, or AWX
- Experience managing PostgreSQL
- Understanding of GitLab CI/CD pipelines
- Written and spoken English
- Background check
Responsibilities
- Build and optimize tooling for monitoring, diagnosis, and incident remediation
- Troubleshoot high-impact production issues
- Participate in an on-call rotation
- Implement and maintain observability solutions
- Contribute to infrastructure lifecycle management
- Apply stability, resiliency, scalability, and security practices
- Maintain technical documentation
- Evolve systems and tools using production feedback
- Collaborate with development teams on infrastructure readiness
- Participate in knowledge-sharing initiatives
Benefits
- Hybrid work with up to 3 remote days per week
- Healthy meals at headquarters and breakfast at all sites
- Swile lunch card for regional-site employees
- Access to a gym, daycare places, and discounted care services
Hiring Process
Recruiter discovery call → manager interview → technical interview → Head of the Tribe interview → HR interview
