Member of Technical Staff - Mid-Training Infra
Reflection is an AI research lab building open frontier models and a full AI stack for developers, enterprises, and public-sector users.
Funding history
About Reflection
Reflection develops open-weight AI models, open-source software for customizing and running agents, AI-factory infrastructure, and related solutions. Its current research emphasizes large language models, reinforcement learning, and agentic reasoning.
Skills
About the Role
You will design, build, and operate GPU infrastructure for high-throughput inference and mid-training workloads. You will develop distributed systems for synthetic data generation and reinforcement learning, optimize model execution and GPU utilization, support large-scale evaluations, and resolve performance bottlenecks across runtimes, kernels, networking, and distributed compute.
Requirements
- GPU infrastructure
- Model serving
- Inference
- GPU optimization
- SGLang
- Megatron
- Reinforcement learning
- Distributed system
- Synthetic data
- GPU kernel
- Networking
Responsibilities
- Design build and operate large-scale GPU infrastructure
- Develop systems for synthetic data generation and reinforcement learning pipelines
- Build high-performance inference platforms across thousands of GPUs
- Optimize inference throughput latency and GPU utilization
- Support distributed reinforcement learning and model evaluation workloads
- Improve model execution through kernel optimization model parallelism and GPU runtime improvements
- Diagnose and resolve performance bottlenecks across distributed systems
Benefits
- Stock options
- Medical insurance
- Dental insurance
- Vision insurance
- Life insurance
- Annual wellness allowance
- Daily office lunch and dinner
- 22 weeks of paid parental leave
- Unlimited paid time off in the U.S.
- 30 days of vacation in the U.K.
- Visa sponsorship
- Regular off-sites
- Happy hours
- Team celebrations
