Software Engineer Platform and Inference
Chai DiscoveryVisit Chai Discovery website
AI molecular-design company building models and a computer-aided design suite for drug discovery.
San Francisco, United States
Funding history
About Chai Discovery
Chai Discovery develops generative AI software that predicts and reprograms biochemical molecular interactions to help scientists design biomolecules and medicines.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own the serving stack for frontier models, improving latency, throughput, GPU efficiency, batching, and autoscaling across a multi-cloud GPU fleet. You will build model pipelines, experiment and observability tooling, and operate 24/7 production systems.
Requirements
- 4+ years building production systems
- Depth in performance, distributed systems, or ML serving
- Experience optimizing model inference
- Experience with GPU utilization, batching, quantization, caching, or kernel-level work
- Experience owning 24/7 systems, including observability, alerting, and incident response
- Experience with 0-to-1 buildouts and 1-to-n scale-ups
Responsibilities
- Own the model-serving stack
- Improve latency, throughput, GPU efficiency, batching, and autoscaling
- Build product-ready model pipelines
- Build experiment and observability tooling
- Operate 24/7 systems with alerting and incident response
- Collaborate with researchers, product engineers, and the commercial team
