Staff Software Engineer, Foundation Model Inference
Databricks is a data and AI platform that lets organizations build analytics, AI agents, and applications on a unified, governed lakehouse.
Maintainer signals as of 9/23/2026
Funding history
Investors
About Databricks
Data engineers, analysts, and AI teams use Databricks to process large datasets, build reliable pipelines, and train models on a single governed platform. Users can run SQL analytics, serve ML predictions in real time, and deploy AI agents grounded in enterprise data. Its open lakehouse architecture provides consistent security and governance across analytical and operational workloads.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will build infrastructure for large-scale LLM inference across partner and self-hosted models. You will help shape the product roadmap, improve distributed workload reliability, latency, and efficiency, and deliver end-to-end AI experiences with platform, infrastructure, and ML stakeholders.
Requirements
- Backend or infrastructure engineering experience
- Distributed systems
- Scalable API
- Cloud-native infrastructure
- Real-time serving
- ML infrastructure
- GPU orchestration
- Service-oriented architecture
- Deployment pipeline
- System observability
- Scala, Go, or Python
Responsibilities
- Build LLM infrastructure for large-scale inference workloads
- Shape the Foundation Model APIs product roadmap and execution
- Improve the reliability, latency, and efficiency of distributed AI workloads
- Collaborate to deliver end-to-end AI experiences
- Shape developer and data scientist interactions with AI products
Benefits
- Annual performance bonus eligibility
- Equity eligibility
