ML Inference Engineer

SpAItial is a London AI company developing physically grounded world models that generate and reason about explorable 3D environments.

London, United Kingdom

Funding history

About SpAItial

SpAItial develops the Echo model family and an app/API for generating, editing, exploring, sharing, and exporting persistent 3D Gaussian Splat worlds from text, images, and panoramas.

View jobs by SpAItial

Skills

About the Role

You will build and operate inference services and APIs from request through generated output. You will optimize GPU utilization, memory use, batching, cold starts, reliability, and cost. You will deploy research models as production services, including customer-specific configurations, and improve inference performance.

Requirements

  • 3+ years of software engineering experience in an AI, machine learning, or computer vision product environment
  • Experience designing and owning production APIs
  • Strong Python skills
  • Experience shipping containerized services on GPU inference platforms
  • Hands-on GPU inference performance experience, including profiling, memory, batching, and cold starts
  • Familiarity with 3D or computer vision data

Responsibilities

  • Own inference services and APIs from request to generated output
  • Optimize serving performance, reliability, cost, GPU utilization, and device memory use
  • Apply inference performance techniques
  • Deploy research models as production services, including customer-specific configurations
  • Bring new models to users