Software Engineer - Distributed Data Systems
Databricks is a data and AI platform that lets organizations build analytics, AI agents, and applications on a unified, governed lakehouse.
Maintainer signals as of 9/25/2026
Funding history
Investors
About Databricks
Data engineers, analysts, and AI teams use Databricks to process large datasets, build reliable pipelines, and train models on a single governed platform. Users can run SQL analytics, serve ML predictions in real time, and deploy AI agents grounded in enterprise data. Its open lakehouse architecture provides consistent security and governance across analytical and operational workloads.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will work on the full development cycle for distributed data systems. You will clarify requirements and make design decisions, produce technical design documents and project plans, develop features, mentor junior engineers, and test, roll out, and monitor production changes.
Requirements
- BS in Computer Science or equivalent practical experience in databases or distributed systems
- Ability to work toward a multi-year vision with incremental deliverables
- Motivation to deliver customer value and impact
- 3+ years of production-level experience with Java, Scala, or C++
- Solid foundation in algorithms and data structures
- Experience with distributed systems, databases, and big data systems including Apache Spark and Hadoop
Responsibilities
- Drive requirements clarity and design decisions for ambiguous problems
- Produce technical design documents and project plans
- Develop new features
- Mentor junior engineers
- Test, roll out, and monitor production changes
