AI Inference Engineer

Baseten is an AI inference platform for deploying, optimizing, and scaling custom, open-source, and fine-tuned models in production.

San Francisco, United States
About Baseten

Baseten provides model runtimes, inference infrastructure, developer workflows, and deployment options including managed cloud, self-hosted, and hybrid environments.

View jobs by Baseten

Skills

About the Role

You will partner directly with customers to architect, build, deploy, and monitor high-scale AI applications. You will translate business goals into reliable services, develop production software, define proofs of concept, optimize AI/ML projects, and own customer projects from exploration through production.

Requirements

  • Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or a related field
  • 2+ years of professional work experience in a fast-paced, high-growth environment
  • Production experience with a general-purpose programming language
  • Familiarity with AI/ML pipelines and the ML model development and deployment lifecycle
  • Communication skills for complex technical topics

Responsibilities

  • Develop and maintain production software systems and product features
  • Design, implement, deploy, and monitor customer solutions
  • Turn objectives into specifications and proofs of concept
  • Optimize and enhance AI/ML projects
  • Own products and customer projects end-to-end
  • Navigate technical tradeoffs and select appropriate tools
  • Take ownership and accountability for your work

Benefits

  • Equity
  • Medical, dental, and vision insurance for U.S. employees and dependents
  • Flexible PTO and company-wide Winter Break
  • Paid parental leave
  • Fertility and family-building stipend through Carrot
  • Company-facilitated 401(k) for U.S. employees
AI Inference Engineer at Baseten | JobStash