Applied Machine Learning Engineer

Fireworks AI operates an AI platform for production inference and training of open-source models.

San Mateo, United States
About Fireworks AI

Fireworks AI provides serverless and dedicated model inference, model deployment, and supervised and reinforcement fine-tuning for developers and enterprises building AI applications.

View jobs by Fireworks AI

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will own customer engagements from scoping through production. You will embed with customer teams, fine-tune and evaluate models, optimize serving performance and cost, and communicate actionable customer needs to product and engineering.

Requirements

  • Demonstrate strength in two or three areas of software engineering, machine learning, and infrastructure
  • Have formal machine learning training and experience post-training LLMs
  • Have experience with GPU infrastructure, distributed serving, or inference engines
  • Have founder or founding engineer experience
  • Have startup or fast-paced environment experience
  • Use Python and at least one systems language
  • Apply AI-assisted and agentic engineering

Responsibilities

  • Own customer deployments from scoping through production
  • Embed with customer teams to understand their domain, data, and constraints
  • Run fine-tuning and post-training work from data through production serving
  • Optimize model selection, deployment shape, inference performance, and cost
  • Assess achievable scope before committing
  • Feed actionable customer needs back to product and engineering
Applied Machine Learning Engineer at Fireworks AI | JobStash