Member of Technical Staff Training
InceptionVisit Inception website
AI research and product company building diffusion-based language models for production applications.
Palo Alto, United States
Funding history
About Inception
Inception develops and deploys the Mercury family of diffusion LLMs, which generate and refine output in parallel to target lower-latency, lower-cost production AI workloads.
Skills
About the Role
You will develop diffusion-language-model architectures, implement training objectives and loss functions, research controlled generation and multimodal methods, improve training and inference efficiency, and develop post-training techniques that improve model behavior.
Requirements
- At least 2 years of experience on ML projects using PyTorch or equivalent
- Familiarity with transformers and core LLM concepts
- Familiarity with diffusion-model training and inference
- Experience training deep learning models at scale in distributed computing environments
Responsibilities
- Design, develop, and optimize diffusion-based language model architectures
- Implement training objectives and loss functions for discrete diffusion LLMs
- Research and implement controlled text generation and constraint-satisfaction techniques
- Develop multimodal integration methods
- Improve model efficiency, training time, and inference throughput
- Develop post-training techniques to align and improve model behavior
Benefits
- Equity
- Flexible vacation and paid time off
- Health, dental, and vision insurance
- 401k match
- Catered meals
- Commuter subsidies
