Staff Forward Deployed Engineer

Tenstorrent is an AI-computing company that sells AI hardware and licenses AI and RISC-V intellectual property.

Toronto, Canada
About Tenstorrent

Tenstorrent builds computers for AI, including AI processors, scalable server systems, and open-source software and compiler tooling. It also licenses AI and RISC-V IP for customers building customized silicon.

View jobs by Tenstorrent

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will work with customers to turn requirements and issues into verifiable acceptance criteria. You will contribute production code, operate AI inference deployments, debug the full inference stack, and provide reproducible code, benchmarks, and telemetry that help improve deployments and products.

Requirements

  • 5+ years of relevant technical experience.
  • Experience translating ambiguous customer requirements or issues into verifiable acceptance criteria.
  • Kubernetes and Helm experience at multi-node, HPC, or AI-cluster scale.
  • Experience with observability and infrastructure automation, including Prometheus, Grafana, or OpenTelemetry.
  • Experience with LLM inference serving engines and technologies, including vLLM, SGLang, Mooncake, NIM, Dynamo, or LMCache.

Responsibilities

  • Work directly with customers to understand challenges and provide effective solutions.
  • Contribute production code and operate deployments.
  • Debug the full inference stack from failing requests through serving layers, memory failures, and kernel dispatch.
  • Translate customer requirements and issues into verifiable acceptance criteria.
  • Provide pull requests, reproducible code, benchmarks, and telemetry data.
Staff Forward Deployed Engineer at Tenstorrent | JobStash