Site Reliability Engineer (Monetization)
Xsolla is a global video game commerce company providing tools and services to launch, monetize, and scale games. Its offerings include payments, web shops, publishing, distribution, LiveOps, anti-fraud, subscriptions, SDKs, and creator solutions for developers, publishers, payment providers, creators, and other gaming businesses.
Maintainer signals as of 8/23/2026
Projects
About Xsolla (USA), Inc.
Xsolla operates as a global merchant of record and video game commerce platform serving developers, publishers, resellers, payment providers, creators, and retailers. It provides payment processing across more than 200 countries and regions, 1,000+ payment methods, and 130+ currencies, alongside tax management, compliance, fraud prevention, refunds, dispute management, and end-user support. Its product portfolio includes Web Shop, Publishing Suite, Payments, Xsolla Pay, Mobile Buy Button, SDKs, Subscriptions, game distribution, Partner Network, Offerwall, LiveOps, Anti-Fraud, Login, Site Builder, cloud gaming, and related gaming commerce tools.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
Own application-level infrastructure and reliability for the Monetization domain, including Kubernetes deployments, observability, CI/CD automation, capacity planning, production readiness, incident investigation, runbooks, and reliability standards.
Requirements
- 3+ years of SRE, DevOps, or platform engineering experience
- Software development background building and shipping backend services
- Kubernetes experience with Helm, manifests, deployment strategies, and application-level debugging
- Observability experience with monitors, dashboards, SLOs, and SLIs
- Infrastructure as Code experience with Terraform or Terragrunt
- GCP experience, including IAM, networking, and managed services
- CI/CD pipeline experience with GitLab CI or GitHub Actions
- Programming or scripting proficiency with Python, Go, or Bash
- Incident response and post-mortem experience
- Collaboration and communication skills
- Experience with payments, fintech, e-commerce, or gaming transactional systems
Responsibilities
- Own application-level infrastructure, including Helm charts, Terraform configurations, Kubernetes deployments, runtime configuration, networking, and integrations
- Design and implement SLOs, SLIs, monitors, alerts, and dashboards for critical services
- Set up and evolve CI/CD pipelines, including deployment and rollback automation
- Perform capacity planning, load testing, performance tuning, and regression investigations
- Run Production Readiness Reviews and define production-readiness standards
- Support incident response, post-mortems, reliability improvements, and runbook maintenance
- Build automation that reduces operational toil
- Maintain a reliability roadmap with product engineering leads
- Participate in product planning, refinements, and architecture reviews
- Co-author company-wide SLO, SLI, capacity, and operational standards
- Participate in the SRE duty rotation
