Distributed Systems Engineer

Beam is an AI-native serverless cloud platform for GPU inference, task queues, and secure sandboxes.

New York City, United States

Funding history

About Beam

Smartshare, Inc. operates Beam, a developer platform for running AI and ML workloads—including inference endpoints, agents, task queues, and sandboxes—on CPUs and GPUs without managing underlying infrastructure.

View jobs by Beam

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build low-level platform systems for container runtimes, OCI image formats, and lazy loading from content-addressable storage. You will improve GPU workload packing across clouds and work with GPU checkpoint restore and CRIU. You will develop backend software primarily in Go, with some Python.

Requirements

  • 3 to 5 years of experience with a large distributed system
  • Knowledge of Kubernetes
  • Experience with a statically typed language such as Go or Rust
  • Familiarity with gRPC, Helm, Kustomize, and Terraform
  • Ability to collaborate closely with customers
  • Knowledge of developer tools, cloud-native technologies, and open-source software

Responsibilities

  • Develop low-level systems for container runtimes and OCI image formats
  • Build lazy loading for large files from content-addressable storage
  • Improve packing of GPU workloads across multiple clouds
  • Work with GPU checkpoint restore and CRIU
  • Develop backend software in Go and Python

Benefits

  • Meaningful equity
  • Health, dental, and vision benefits with 90% coverage for employees and 50% for dependents
  • Opportunities to participate in cloud-native community events
  • Fitness stipend
  • Learning budget
Distributed Systems Engineer at Beam | JobStash