SRE/DevOps Engineer
KodyPay, also known as Kody, offers an all-in-one platform to simplify in-person commerce for businesses. It provides a suite of products including intuitive payment solutions, instant access to funds with the KodyCard, and powerful management and analytics tools. Partnering with Adyen N.V., Kody enables businesses to accept a wide range of payment methods, streamlining operations for both merchants and customers.
Investors
Projects
About KodyPay
KodyPay Ltd, registered in England and Wales, provides a comprehensive, all-in-one platform designed to simplify in-person commerce for businesses, particularly in the hospitality and retail sectors. The company's offerings include a suite of intuitive payment solutions such as terminals and online ordering, instant cash access via the KodyCard, and robust management tools with built-in loyalty and analytics features for understanding customer behavior. Through a partnership with financial services provider Adyen N.V., Kody supports a wide array of payment methods, including major credit cards and digital wallets like Apple Pay and Google Pay. The platform aims to enhance profit margins by improving the efficiency of the payment ecosystem, while also emphasizing personal customer service and handling processes like payment management, returns, and Know Your Customer (KYC) checks.
Skills
About the Role
You will ensure production reliability and runbooks, partnering with the Platform Engineering team to implement reliability guardrails to meet uptime and SLA requirements on AWS. You will own deployment pipelines and code management on GitHub. You will lead rapid-response troubleshooting during production incidents and conduct blameless postmortems to improve systems. You will implement advanced monitoring, logging, and alerting to proactively detect and mitigate issues. You will act as a technical bridge between US operations and international engineering hubs, coordinating across time zones with bilingual communication.
Requirements
- Experience deploying and operating applications on AWS with mastery of GitHub workflows and actions.
- Strong monitoring tools, log management, and scripting skills for quick triaging and troubleshooting.
- Ownership and transparency; highly responsive and communicative with end-to-end responsibility for production health.
- High psychological resilience; able to stay positive and focused during high-stakes incidents.
- Fluency in Mandarin and English for effective cross-border collaboration.
Responsibilities
- Partner with the Platform Engineering team to implement reliability guardrails, ensuring applications running on AWS meet uptime and SLA requirements.
- Own the deployment pipelines and code management practices extensively via GitHub.
- Lead rapid-response troubleshooting during production incidents; conduct blameless post-mortems to continuously harden our systems.
- Implement advanced monitoring, logging, and alerting systems to proactively detect and mitigate system anomalies.
- Act as a key technical bridge between US operations and international engineering hubs, leveraging bilingual communication to streamline complex technical alignment.
Benefits
- Competitive packages aligned with California market standards
- Lead a dynamic and innovative team in a very rapidly growing company
- Collaborative, inclusive environment where your contributions are recognized and valued
