Research Intern Inference Winter 2027
Together AIVisit Together AI website
Together AI operates an AI-native cloud platform for open and custom AI models.
San Francisco, United States
About Together AI
Together AI provides production AI infrastructure spanning inference, accelerated compute, model training and fine-tuning, and secure code sandboxes for AI development.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will investigate distributed inference, compiler-aware optimization, and inference-time computation strategies. You will co-design and implement optimizations across models, systems, and hardware, conduct rigorous experiments, communicate project results, and document findings in publications and blog posts.
Requirements
- Currently pursuing a final-year Bachelor's, Master's, or Ph.D. degree in Computer Science, Electrical Engineering, or a related field
- Strong knowledge of Machine Learning and Deep Learning fundamentals
- Experience with deep learning frameworks such as PyTorch and JAX
- Strong programming skills in Python
- Familiarity with Transformer architectures and recent developments in foundation models
Responsibilities
- Design and conduct rigorous experiments to validate hypotheses
- Communicate project plans, progress, and results to the broader team
- Document findings in scientific publications and blog posts
Benefits
- Housing stipends
- Other competitive benefits
