Member of Technical Staff - Compute Platform
Prime Intellect builds the 'Open Superintelligence Stack' — an integrated compute, training, inference, and sandbox platform that lets companies train, deploy, and continuously improve their own AI models and agents. It serves AI startups, 'neolabs', and enterprises (over 6,000 customers, including Ramp and Zapier) that want to own their model optimization loop rather than rely solely on closed frontier labs.
Maintainer signals as of 8/12/2026
Funding history
Investors
Projects
About Prime Intellect
Prime Intellect is a San Francisco-based AI infrastructure company building what it calls the Open Superintelligence Stack: a full-stack platform spanning GPU compute (on-demand and reserved clusters), large-scale reinforcement learning training ('Lab'), an Environments Hub with 2,500+ community RL environments, hosted evaluations, sandboxed code execution, and dedicated/serverless model inference with native LoRA support. The company maintains open-source libraries (verifiers and prime-rl) used to build and train RL environments, and publishes frontier open research such as the INTELLECT and SYNTHETIC model/dataset series. Prime Intellect works with AI startups, enterprises, and 'neolab' customers such as Ramp and Zapier, helping them turn production traces and evaluations into custom-trained, post-trained agent models that outperform closed frontier models on specific workflows at lower cost and latency. The company has raised over $150M in total funding, including a $130M Series A led by Radical Ventures with participation from NVIDIA Ventures, Intel Capital, and Dell Technologies Capital, and reports over $100M in annualized revenue.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will build the platform software and infrastructure used to manage and monitor AI workloads. You will develop web interfaces, Python APIs and backend services, real-time debugging tools, distributed training infrastructure in Rust, automation pipelines, cloud resources, container orchestration, and scheduling systems for heterogeneous hardware.
Requirements
- Strong Python backend development with FastAPI and async
- Modern frontend development with TypeScript, React/Next.js, and Tailwind
- Developer tools and dashboard development
- RESTful API design and implementation
- Rust systems programming
- Ansible and Terraform automation
- Kubernetes
- GCP or other cloud platforms
- Prometheus and Grafana
- GPU computing or ML infrastructure experience
Responsibilities
- Build web interfaces for AI workload management and monitoring
- Develop REST APIs and backend services in Python
- Create real-time monitoring and debugging tools
- Implement resource management and job control features
- Design distributed training infrastructure in Rust
- Build networking and coordination components
- Create Ansible infrastructure automation pipelines
- Manage cloud resources and container orchestration
- Implement scheduling systems for CPU, GPU, and TPU hardware
- Integrate backend features into existing infrastructure
