Software Engineer Platform and Inference

AI molecular-design company building models and a computer-aided design suite for drug discovery.

San Francisco, United States
About Chai Discovery

Chai Discovery develops generative AI software that predicts and reprograms biochemical molecular interactions to help scientists design biomolecules and medicines.

View jobs by Chai Discovery

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will own the serving stack for frontier models, improving latency, throughput, GPU efficiency, batching, and autoscaling across a multi-cloud GPU fleet. You will build model pipelines, experiment and observability tooling, and operate 24/7 production systems.

Requirements

  • 4+ years building production systems
  • Depth in performance, distributed systems, or ML serving
  • Experience optimizing model inference
  • Experience with GPU utilization, batching, quantization, caching, or kernel-level work
  • Experience owning 24/7 systems, including observability, alerting, and incident response
  • Experience with 0-to-1 buildouts and 1-to-n scale-ups

Responsibilities

  • Own the model-serving stack
  • Improve latency, throughput, GPU efficiency, batching, and autoscaling
  • Build product-ready model pipelines
  • Build experiment and observability tooling
  • Operate 24/7 systems with alerting and incident response
  • Collaborate with researchers, product engineers, and the commercial team
Software Engineer Platform and Inference at Chai Discovery | JobStash