Software Engineer Distributed Systems

fal is an active generative-media AI platform for developers, providing optimized model APIs, serverless deployment, and GPU compute.

San Francisco, United States
About fal

Founded in 2021 by Burkay Gur and Gorkem Yurtseven, fal provides infrastructure for production generative-media applications, including image, video, audio, 3D, and multimodal models.

View jobs by fal

Skills

About the Role

You will build core platform services for request routing, AI workload orchestration, scheduling, GPU autoscaling, file storage, and queueing. You will design for major traffic growth, automate routine work with AI, and profile and tune CPU and memory performance.

Requirements

  • Distributed systems
  • Python
  • Rust
  • consensus
  • scheduling
  • fault tolerance
  • capacity planning
  • computational complexity
  • memory allocation
  • observability
  • AI inference
  • AI training
  • systems programming
  • networking
  • GPU

Responsibilities

  • Build platform services for request routing, orchestration, scheduling, autoscaling, storage, and queueing
  • Design platform evolution for increased global traffic and low latency
  • Automate routine systems-development work with AI
  • Profile and tune low-level CPU and memory performance

Benefits

  • Regular team events and offsites
Software Engineer Distributed Systems at fal | JobStash