Search...

Member of Technical Staff - Compute Platform

Prime Intellect logo
Prime Intellect

Prime Intellect builds the 'Open Superintelligence Stack' — an integrated compute, training, inference, and sandbox platform that lets companies train, deploy, and continuously improve their own AI models and agents. It serves AI startups, 'neolabs', and enterprises (over 6,000 customers, including Ramp and Zapier) that want to own their model optimization loop rather than rely solely on closed frontier labs.

Seed33 current maintainers31 active leads9 new active leads3 lead step-downsTeam intelligence

Maintainer signals as of 8/12/2026

San Francisco, USA
About Prime Intellect

Prime Intellect is a San Francisco-based AI infrastructure company building what it calls the Open Superintelligence Stack: a full-stack platform spanning GPU compute (on-demand and reserved clusters), large-scale reinforcement learning training ('Lab'), an Environments Hub with 2,500+ community RL environments, hosted evaluations, sandboxed code execution, and dedicated/serverless model inference with native LoRA support. The company maintains open-source libraries (verifiers and prime-rl) used to build and train RL environments, and publishes frontier open research such as the INTELLECT and SYNTHETIC model/dataset series. Prime Intellect works with AI startups, enterprises, and 'neolab' customers such as Ramp and Zapier, helping them turn production traces and evaluations into custom-trained, post-trained agent models that outperform closed frontier models on specific workflows at lower cost and latency. The company has raised over $150M in total funding, including a $130M Series A led by Radical Ventures with participation from NVIDIA Ventures, Intel Capital, and Dell Technologies Capital, and reports over $100M in annualized revenue.

View jobs by Prime Intellect

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build the platform software and infrastructure used to manage and monitor AI workloads. You will develop web interfaces, Python APIs and backend services, real-time debugging tools, distributed training infrastructure in Rust, automation pipelines, cloud resources, container orchestration, and scheduling systems for heterogeneous hardware.

Requirements

  • Strong Python backend development with FastAPI and async
  • Modern frontend development with TypeScript, React/Next.js, and Tailwind
  • Developer tools and dashboard development
  • RESTful API design and implementation
  • Rust systems programming
  • Ansible and Terraform automation
  • Kubernetes
  • GCP or other cloud platforms
  • Prometheus and Grafana
  • GPU computing or ML infrastructure experience

Responsibilities

  • Build web interfaces for AI workload management and monitoring
  • Develop REST APIs and backend services in Python
  • Create real-time monitoring and debugging tools
  • Implement resource management and job control features
  • Design distributed training infrastructure in Rust
  • Build networking and coordination components
  • Create Ansible infrastructure automation pipelines
  • Manage cloud resources and container orchestration
  • Implement scheduling systems for CPU, GPU, and TPU hardware
  • Integrate backend features into existing infrastructure