Staff Software Engineer Foundation Model Inference

12 hours agoLeadSalary: 190K - 265KSan Francisco, USAAiJobs by Databricks

Databricks is a data and AI platform that lets organizations build analytics, AI agents, and applications on a unified, governed lakehouse.

160 Spear Street, Suite 1300, San Francisco, CA 94105, United States
About Databricks

Data engineers, analysts, and AI teams use Databricks to process large datasets, build reliable pipelines, and train models on a single governed platform. Users can run SQL analytics, serve ML predictions in real time, and deploy AI agents grounded in enterprise data. Its open lakehouse architecture provides consistent security and governance across analytical and operational workloads.

View jobs by Databricks

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build infrastructure for large-scale LLM inference across partner and self-hosted models. You will improve reliability, latency, and efficiency of distributed AI workloads, collaborate on end-to-end platform experiences, and shape AI developer workflows.

Requirements

  • 8+ years of backend or infrastructure engineering experience
  • Experience with distributed systems, scalable APIs, or cloud-native infrastructure
  • Experience with real-time serving, ML infrastructure, or GPU orchestration
  • Familiarity with service-oriented architecture, deployment pipelines, and system observability

Responsibilities

  • Build LLM infrastructure for large-scale customer inference workloads
  • Improve reliability, latency, and efficiency of distributed AI workloads
  • Collaborate with platform, infrastructure, and ML teams on end-to-end experiences
  • Shape developer and data-scientist AI workflows

Benefits

  • Annual performance bonus
  • Equity