Staff Software Engineer Distributed Data Systems

Databricks is a data and AI platform that lets organizations build analytics, AI agents, and applications on a unified, governed lakehouse.

Series F+0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/23/2026

160 Spear Street, Suite 1300, San Francisco, CA 94105, United States
About Databricks

Data engineers, analysts, and AI teams use Databricks to process large datasets, build reliable pipelines, and train models on a single governed platform. Users can run SQL analytics, serve ML predictions in real time, and deploy AI agents grounded in enterprise data. Its open lakehouse architecture provides consistent security and governance across analytical and operational workloads.

View jobs by Databricks

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build next-generation distributed data storage and processing systems for workloads ranging from ETL to data science. You may develop Apache Spark, data-plane storage services and client libraries, Delta Lake capabilities, data-pipeline orchestration, or query optimization and execution technology.

Requirements

  • BS or higher in Computer Science, a related technical field, or equivalent practical experience.
  • Work toward a multi-year vision with incremental deliverables.
  • Deliver customer value and impact.
  • Have 8+ years of production-level experience in Java, Scala, or C++.
  • Apply algorithms and data structures to real-world use cases.
  • Have experience with distributed systems, databases, and big data systems such as Apache Spark and Hadoop.

Responsibilities

  • Build distributed data storage and processing systems.
  • Develop data systems for ETL and data science workloads.
  • Build high-performance query optimization and execution technology.

Benefits

  • Eligibility for an annual performance bonus.
  • Equity.