Early Career Software Engineer

AI inference cloud providing OpenAI-compatible APIs, private deployments, and GPU infrastructure for production AI workloads.

Palo Alto, United States
About DeepInfra

DeepInfra operates an AI inference cloud for running LLMs, vision, embeddings, image/video generation, speech, and other machine-learning models at scale, including private GPU deployments and GPU rental.

View jobs by DeepInfra

Skills

About the Role

You will collaborate on inference solutions for AI models, implement and optimize models, and monitor production systems. You will build features, fix bugs, contribute to code reviews and design discussions, and experiment with AI and ML techniques that improve model performance.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field, completed or in the final year
  • Data structures, algorithms, and software design fundamentals
  • Proficiency in Python
  • Experience with AI and ML libraries and frameworks including NumPy, pandas, SciPy, TensorFlow, or PyTorch
  • AI and ML experience through coursework, research, projects, employment, or internships
  • Familiarity with AI models, Transformers, and Diffusers
  • Experience with Git and agile development methodologies
  • Problem-solving ability, including debugging and code optimization
  • Communication and collaboration skills

Responsibilities

  • Collaborate to design, develop, and test inference solutions for AI models
  • Implement, optimize, and evaluate AI models using Python, C++, CUDA, and NCCL
  • Monitor and maintain production model-serving systems
  • Build features and fix bugs
  • Contribute to code reviews
  • Participate in standups, design reviews, and team discussions
  • Explore AI and ML techniques and tools
  • Experiment with improving model performance
Early Career Software Engineer at DeepInfra | JobStash