Staff Software Engineer, Infrastructure
Ripple provides payments, custody and stablecoin solutions that help financial institutions integrate blockchain and digital assets.
Funding history
Investors
Projects
About Ripple
Ripple helps financial institutions transform global payments by providing blockchain-powered infrastructure for cross-border payments, digital asset custody, and stablecoin solutions. With it, users can enable instant settlements, reduce costs, and access new markets. The company was originally founded as OpenCoin in 2012 and rebranded to Ripple in 2015.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You'll play a key role in shaping engineering practices, ensuring the reliability and performance of Ripple's platform across a multi-cloud environment, and empowering the engineering team to deliver innovative features with speed and efficiency. You will drive improvements in automation, observability, and overall platform stability.
Requirements
- 10+ years of experience in software engineering, platform engineering, infrastructure engineering, or systems operations for highly available production systems.
- Experience managing blockchain nodes, including high availability, failover, and operational resilience.
- Experience operating infrastructure in high-traffic, customer-critical, or security-sensitive environments.
- Strong production experience with PostgreSQL.
- Experience provisioning and managing infrastructure with Terraform or similar infrastructure-as-code tools.
- Deep experience with containerized infrastructure and Kubernetes in highly available environments.
- Proficiency with .NET, Go, Bash, or TypeScript.
- Experience with RabbitMQ or AMQP.
- Experience with OpenTelemetry, Grafana, Loki, Prometheus, or similar observability platforms.
- Familiarity with GitOps practices using Argo CD or Flux.
- Experience with confidential computing solutions such as AWS Nitro Enclaves, IBM Hyper Protect Virtual Servers, or GCP Confidential Computing.
- Strong problem-solving, distributed-systems diagnosis, communication, documentation, and operational-excellence skills.
Responsibilities
- Design, build, and operate scalable, resilient infrastructure across Azure, AWS, GCP, and IBM Cloud.
- Lead infrastructure architecture and reliability improvements for critical custody services.
- Implement and improve monitoring, alerting, logging, and observability across distributed systems.
- Own and evolve blockchain node infrastructure, including high availability, failover, and provider management.
- Build automation for infrastructure provisioning, deployments, testing, failover, incident response, and operational maintenance.
- Drive deployment operations and release reliability for platform and product services.
- Proactively identify performance bottlenecks, reliability risks, and operational gaps before they impact customers.
- Participate in on-call rotations, support production incidents, and improve incident response and post-incident learning.
- Contribute to internal platform tools, services, and developer workflows.
- Create documentation, runbooks, and operational procedures for critical systems.
- Mentor engineers and provide technical guidance on infrastructure, reliability, and platform engineering decisions.
Benefits
- Professional development budget
- Flexible in-office collaboration days
- Bi-weekly all-company meeting with the Leadership Team
- Team offsites, team bonding activities, and happy hours
- Competitive bonuses and equity
- Health, retirement, family forming, and family support benefits
- Employee giving match
- Mobile phone stipend
- R&R days
- Wellness reimbursement and weekly onsite and virtual programming
- Generous vacation policy
- Parental leave and family planning benefits
- Catered lunches and fully stocked kitchens
