Senior Site Reliability Engineer - GPU Fleet
Skills
About the Role
You will help build and operate Sesterce's sovereign AI infrastructure for production customers across Europe. You will own high-impact work across GPU fleet reliability, observability, incident response, and automation. You will partner with engineering, infrastructure, operations, and customer teams to improve reliability, delivery speed, and operational clarity across Sesterce sites.
Requirements
- Strong experience in site reliability engineering or an adjacent technical field.
- Clear written communication and comfort working with infrastructure teams.
- Pragmatic judgment in production environments where reliability matters.
Responsibilities
- Own high-impact work across GPU fleet reliability, observability, incident response, and automation.
- Partner with engineering, infrastructure, operations, and customer teams.
- Improve reliability, delivery speed, and operational clarity across Sesterce sites.
