Member of Technical Staff Engineering Lead Compute Platform
ReflectionVisit Reflection website
Reflection is an AI research lab building open frontier models and a full AI stack for developers, enterprises, and public-sector users.
New York, United States
Funding history
About Reflection
Reflection develops open-weight AI models, open-source software for customizing and running agents, AI-factory infrastructure, and related solutions. Its current research emphasizes large language models, reinforcement learning, and agentic reasoning.
Skills
About the Role
You will lead systems engineers building and operating a Kubernetes-based multi-cloud compute platform. You will prioritize delivery, guide architecture, contribute hands-on where needed, work with training teams on reliability, and manage critical vendor relationships.
Requirements
- Experience building and growing systems or infrastructure teams
- Systems-level engineering experience
- Strong coding ability
- Knowledge of orchestration, storage, or GPU hardware
- Kubernetes-first architecture experience
- Vendor management experience
- Ability to guide strategy across a multi-cloud environment
Responsibilities
- Build, mentor, and grow a systems engineering team
- Lead compute fleet reliability and availability efforts
- Prioritize work and manage projects
- Guide scalable and reliable platform architecture
- Contribute targeted hands-on engineering work
- Co-design fault tolerance and remediation with training teams
- Manage vendor relationships and deals
- Prepare the fleet for next-generation GPUs and larger clusters
Benefits
- Stock options
- Medical, dental, vision, and life insurance
- Annual wellness allowance
- Daily office lunch and dinner
- 22 weeks of paid parental leave
- Unlimited paid time off in the U.S.
- 30 days of vacation in the U.K.
- Visa sponsorship support
- Regular off-sites, happy hours, and team celebrations
