AI Inference Engineer
1 hour agoSeniorSalary: 165K - 330KSan Francisco; Montreal; New York; Remote; TorontoHybridFull TimeForward Deployed EngineerJobs by Baseten
BasetenVisit Baseten website
Baseten is an AI inference platform for deploying, optimizing, and scaling custom, open-source, and fine-tuned models in production.
San Francisco, United States
Funding history
About Baseten
Baseten provides model runtimes, inference infrastructure, developer workflows, and deployment options including managed cloud, self-hosted, and hybrid environments.
Skills
About the Role
You will partner directly with customers to architect, build, deploy, and monitor high-scale AI applications. You will translate business goals into reliable services, develop production software, define proofs of concept, optimize AI/ML projects, and own customer projects from exploration through production.
Requirements
- Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or a related field
- 2+ years of professional work experience in a fast-paced, high-growth environment
- Production experience with a general-purpose programming language
- Familiarity with AI/ML pipelines and the ML model development and deployment lifecycle
- Communication skills for complex technical topics
Responsibilities
- Develop and maintain production software systems and product features
- Design, implement, deploy, and monitor customer solutions
- Turn objectives into specifications and proofs of concept
- Optimize and enhance AI/ML projects
- Own products and customer projects end-to-end
- Navigate technical tradeoffs and select appropriate tools
- Take ownership and accountability for your work
Benefits
- Equity
- Medical, dental, and vision insurance for U.S. employees and dependents
- Flexible PTO and company-wide Winter Break
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k) for U.S. employees
