Research Tinker RL Systems
Thinking Machines LabVisit Thinking Machines Lab website
Artificial-intelligence research and product company building customizable AI systems, including the Tinker training API and Inkling open-weight models.
Distributed
Funding history
About Thinking Machines Lab
Thinking Machines Lab develops AI products that let researchers and developers fine-tune and use models, while also releasing open-weight multimodal models.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will develop model-customization techniques and post-training systems, spanning training recipes, data pipelines, numerics, and kernels. You will co-design RL algorithms and systems, debug production RL runs, optimize post-training pipelines, work with users, and help them achieve frontier-level results.
Requirements
- Bachelor’s degree or equivalent experience in a relevant discipline
- Proficiency in Python
- Familiarity with PyTorch, TensorFlow, or JAX
- Ability to debug distributed training and write scalable code
- Written technical communication
- Interest in Tinker and its adoption
- Probability, statistics, and machine learning fundamentals
- Reinforcement-learning training stability
- Low-precision training and inference
- Quantization
- LLM serving
- Scaling studies
Responsibilities
- Develop frontier model-customization techniques and post-training systems
- Co-design reinforcement-learning algorithms and training systems across the stack
- Debug reinforcement-learning runs
- Optimize post-training pipelines
- Work directly with researchers and companies using Tinker
- Help users achieve frontier-level results
Benefits
- Health, dental, and vision benefits
- Unlimited PTO
- Paid parental leave
- Relocation support as needed
