Search...

Staff Machine Learning Engineer

webAI logo
webAI

webAI is a sovereign AI platform that lets organizations build, deploy, and operate custom AI on their own infrastructure. It serves enterprise teams across industries such as healthcare, manufacturing, financial services, aviation, public sector, and retail.

Austin, USA
About webAI

webAI provides a full-stack, locally operated AI platform for enterprises seeking private, controllable AI. Its platform includes Navigator for building, training, and deploying custom models; Companion, an on-device AI assistant; Runtime for distributed workload orchestration; webFrame for optimized inference and training; Network for secure local connectivity; and a CLI for programmatic management. The company emphasizes data sovereignty, local deployment, predictable costs, and support for edge devices and private clusters.

View jobs by webAI

Skills

About the Role

You will lead the development and optimization of Large Language Models and Mixture of Experts models. You will deliver product initiatives involving on-device inference optimization, quantization, retrieval-augmented generation, agentic frameworks, and tool calling. You will work cross-functionally to integrate models into the platform, conduct research to improve model performance and efficiency, and apply advances in AI and machine learning. You will lead complex projects, mentor junior engineers, share knowledge and best practices, and lead a sub-team.

Requirements

  • Advanced degree in Computer Science or a related field; Ph.D. preferred
  • Proven record of building and innovation through publications or industry experience
  • At least 6 years of machine learning experience with expertise in Large Language Models and Mixture of Experts
  • Strong Python programming skills
  • Experience with TensorFlow and/or PyTorch
  • Ability to lead complex projects and collaborate in a team environment

Responsibilities

  • Lead the development and optimization of Large Language Models and Mixture of Experts models
  • Collaborate with cross-functional teams to integrate machine learning models into the platform
  • Conduct machine learning research to improve the performance and efficiency of Large Language Models
  • Apply advances in AI and machine learning to improve models and methodologies
  • Mentor junior engineers and contribute to knowledge sharing and best practices
  • Lead core product initiatives across on-device inference optimization, quantization, retrieval-augmented generation, agentic frameworks, and tool calling
  • Deliver projects end-to-end with other engineering functions
  • Lead the sub-team

Benefits

  • Comprehensive health, dental, and vision benefits package
  • 401(k) match
  • Equity options
  • $200/month Health & Wellness stipend
  • Continuing Education support
  • $500/year Function Health subscription
  • Free parking for in-office employees
  • Flexible Time Off
  • Parental leave for eligible employees
  • Supplemental life insurance