Software Engineer, Distributed Systems

fal is an active generative-media AI platform for developers, providing optimized model APIs, serverless deployment, and GPU compute.

San Francisco, United States
About fal

Founded in 2021 by Burkay Gur and Gorkem Yurtseven, fal provides infrastructure for production generative-media applications, including image, video, audio, 3D, and multimodal models.

View jobs by fal

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build core Python and Rust platform capabilities for request routing, AI workload orchestration, scheduling, GPU autoscaling, storage, and queueing. You will design the platform’s evolution for substantially higher traffic and global low latency, automate routine systems work with AI, and profile and tune CPU and memory performance.

Requirements

  • 3+ years of experience building distributed compute and orchestration platforms in Python or Rust
  • Understanding of distributed systems fundamentals, including consensus, scheduling, fault tolerance, and capacity planning
  • Understanding of computational complexity and memory allocation
  • Track record of designing systems that scale under production load
  • Experience using observability to guide performance and reliability decisions
  • Communication skills and ability to drive technical decisions across teams
  • Experience with AI or ML inference or training infrastructure
  • Experience with high-performance systems programming
  • Background in multi-tenant compute platforms
  • Understanding of networking fundamentals and performance characteristics
  • Familiarity with GPU workload characteristics and scheduling constraints

Responsibilities

  • Build Python and Rust platform services for routing, orchestration, scheduling, autoscaling, storage, and queueing
  • Design platform evolution for increased traffic and global low latency
  • Automate systems work using AI
  • Profile and tune low-level CPU and memory performance

Benefits

  • Equity
  • Relocation assistance to San Francisco
  • Health, dental, and vision insurance
  • Regular team events and offsites
Software Engineer, Distributed Systems at fal | JobStash