Observability Specialist
Wave Mobile Money is a financial technology company providing affordable mobile money services across Africa. Its offerings include money transfers, deposits and withdrawals, bill payments, airtime purchases, merchant payments, collections, bulk payments, and checkout APIs.
Funding history
About Wave Mobile Money Inc.
Wave Mobile Money builds and operates a mobile financial network intended to help make Africa the first cashless continent. Its consumer services enable users to deposit and withdraw money, send funds, pay bills, buy airtime, and access customer support through mobile applications and agents. Wave Business provides bulk payments, in-person merchant payments, cash collection tools, a business portal, and checkout APIs. The company serves consumers, agents, merchants, businesses, and other clients across multiple African countries.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own and evolve the observability platform across backend services, APIs, databases, Kubernetes workloads, cloud infrastructure, and on-premises environments. You will improve system instrumentation, define meaningful service-level indicators, reduce alert fatigue, build internal tooling, investigate incidents, analyse performance, establish observability standards, and manage observability costs as data volumes grow.
Requirements
- 5+ years of experience in observability SRE platform engineering infrastructure engineering backend engineering or production systems engineering
- Deep understanding of metrics logging tracing profiling alerting dashboards service-level indicators and incident response workflows
- Experience building internal tools libraries automation or platforms for engineers
- Excellent communication and collaboration skills
- Pragmatic judgment about tooling improvements and simplification
- Experience working with other people's code
- Proficiency in at least one backend language preferably Python
- Experience with observability tools such as Prometheus Grafana Datadog OpenTelemetry Jaeger Tempo Loki Honeycomb or Sentry
- Experience with Postgres CockroachDB Redis GraphQL or Kubernetes
- Experience with OpenTelemetry instrumentation and collector configuration at scale
Responsibilities
- Improve understanding of production behaviour across applications APIs databases caches workloads and infrastructure
- Define meaningful service-level indicators
- Reduce alert fatigue and make alerts actionable
- Build internal tooling and self-service workflows
- Help engineers instrument services and investigate incidents
- Analyse performance and system dependencies
- Identify reliability latency capacity and cost issues
- Establish observability standards documentation and training materials
- Operate and improve the observability platform
- Control observability costs as data volume grows
- Manage and evolve observability tools
- Support teams in implementing service-level objectives
- Improve alert quality across engineering teams
- Maintain observability modules and code
