Senior Software Engineer in Data
Kamu provides an open-source protocol, the Open Data Fabric (ODF), and a suite of tools for decentralized data exchange and transformation.
Maintainer signals as of 9/2/2026
About Kamu Data Inc.
Kamu is building a decentralized data processing network based on its Open Data Fabric (ODF) protocol, designed for the decentralized exchange and transformation of semi-structured data. It treats data as dynamic streams with a complete, immutable history, enabling verifiable trust and data provenance through a 'Data-as-Code' approach using a novel stream-processing SQL. This creates auditable and autonomous data pipelines, allowing organizations to build data supply chains similar to open-source software development. The Kamu ecosystem includes a command-line interface (CLI) for data management, a scalable Node for operating data pipelines, and a Web Platform to explore the data network. The project's technology and protocols are developed openly with community-governed specifications to foster collaboration and interoperability.
Skills
About the Role
You will evolve core data formats and protocols, improve data engines, and build distributed infrastructure for data pipelines and API queries. You will design data access APIs, build federated data-sharing and compute capabilities, integrate third-party data and blockchain technologies, research privacy-preserving and provenance features, and contribute to documentation and automated testing.
Requirements
- BSc in computer science or equivalent experience
- 6+ years of industry experience
- High proficiency in Rust Java or Scala
- Strong knowledge of SQL and database internals
- Knowledge of modern data lakehouse architecture and horizontal scaling
- Experience with data science toolkits such as Pandas and R
- Knowledge of data integration systems and patterns
- Knowledge of software quality test pyramids and CI/CD
- Structured data format experience with Parquet and Arrow preferred
- Stateful stream-processing fundamentals preferred
- CDC and event sourcing experience preferred
- Docker AWS and Kubernetes experience preferred
- Data visualization experience with PowerBI Tableau or Jupyter preferred
- Agile and Scrum experience preferred
- Open source collaboration experience preferred
- Blockchain indexing and analytics experience preferred
- Decentralized storage experience with IPFS preferred
- Good written English and ability to write clear documentation
Responsibilities
- Evolve core data formats and protocols
- Improve existing data engines and integrate new engines
- Build distributed processing infrastructure for data pipelines and API queries
- Design data access APIs for data ingress and egress
- Build a federated data-sharing and compute network
- Integrate third-party data providers and consumers
- Integrate blockchain decoding and indexing technologies
- Research and implement privacy-preserving compute fine-grained provenance and AI/ML integration
- Communicate technical progress to users and the community
- Contribute to product documentation and automated testing
Benefits
- Remote work with flexible hours
- Equity compensation
- $1,500 home office equipment stipend
- 21 days of paid vacation per year
- Conference travel and education budget
- Home office equipment support for Ukrainian applicants
- Relocation support for Ukrainian applicants
