Research Tinker RL Systems

Artificial-intelligence research and product company building customizable AI systems, including the Tinker training API and Inkling open-weight models.

Distributed
About Thinking Machines Lab

Thinking Machines Lab develops AI products that let researchers and developers fine-tune and use models, while also releasing open-weight multimodal models.

View jobs by Thinking Machines Lab

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will develop model-customization techniques and post-training systems, spanning training recipes, data pipelines, numerics, and kernels. You will co-design RL algorithms and systems, debug production RL runs, optimize post-training pipelines, work with users, and help them achieve frontier-level results.

Requirements

  • Bachelor’s degree or equivalent experience in a relevant discipline
  • Proficiency in Python
  • Familiarity with PyTorch, TensorFlow, or JAX
  • Ability to debug distributed training and write scalable code
  • Written technical communication
  • Interest in Tinker and its adoption
  • Probability, statistics, and machine learning fundamentals
  • Reinforcement-learning training stability
  • Low-precision training and inference
  • Quantization
  • LLM serving
  • Scaling studies

Responsibilities

  • Develop frontier model-customization techniques and post-training systems
  • Co-design reinforcement-learning algorithms and training systems across the stack
  • Debug reinforcement-learning runs
  • Optimize post-training pipelines
  • Work directly with researchers and companies using Tinker
  • Help users achieve frontier-level results

Benefits

  • Health, dental, and vision benefits
  • Unlimited PTO
  • Paid parental leave
  • Relocation support as needed