Software Engineer Intern
DeepInfraVisit DeepInfra website
AI inference cloud providing OpenAI-compatible APIs, private deployments, and GPU infrastructure for production AI workloads.
Palo Alto, United States
Funding history
About DeepInfra
DeepInfra operates an AI inference cloud for running LLMs, vision, embeddings, image/video generation, speech, and other machine-learning models at scale, including private GPU deployments and GPU rental.
Skills
About the Role
You will collaborate with engineers to design, develop, and test inference solutions for AI models. You will implement and optimize models, monitor and maintain the live service, develop features, fix bugs, review code, and participate in stand-ups and design discussions.
Requirements
- Pursue a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
- Demonstrate knowledge of data structures, algorithms, and software design patterns
- Demonstrate proficiency in Python and AI/ML libraries or frameworks
- Demonstrate familiarity with AI models, Transformers, and Diffusers
- Demonstrate experience with Git and agile development methodologies
- Demonstrate problem-solving, debugging, optimization, communication, and teamwork skills
Responsibilities
- Collaborate with engineers to design, develop, and test AI-model inference solutions
- Implement and optimize AI models using Python, C++, CUDA, and NCCL
- Monitor and maintain the live service
- Develop features, fix bugs, and contribute to code reviews
- Participate in stand-ups, code reviews, and design discussions
- Stay current with AI and machine-learning developments
