Solutions Architect US
FuriosaAI is a South Korean AI semiconductor company building energy-efficient inference accelerators, servers, and software for enterprise and cloud AI deployments.
Funding history
Investors
About FuriosaAI
FuriosaAI develops the RNGD AI inference accelerator and NXT RNGD Server, alongside a software toolchain for compiling, optimizing, and deploying LLM and agentic-AI workloads.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will enable US customers to deploy AI models on RNGD NPUs using the Furiosa SDK. You will develop proofs of concept, run benchmarks and debugging sessions, support pre-sales evaluations, train customers, present at technical forums, and relay customer feedback to product and engineering teams.
Requirements
- 2–5 years in a US customer-facing technical role
- Knowledge of the AI and LLM landscape, inference frameworks, and serving stacks
- Experience with vLLM, SGLang, TensorRT-LLM, Triton Inference Server, or similar tools
- Experience with LangChain, LlamaIndex, LangGraph, AutoGen, or MCP-based tooling
- Proficiency in Python
- Knowledge of PyTorch or TensorFlow
- Authorization to work in the US
- Ability to travel to customer sites and Seoul headquarters periodically
Responsibilities
- Own end-to-end technical enablement for US customers deploying AI models on RNGD NPUs
- Develop proofs of concept, benchmarking studies, and live debugging sessions in customer environments
- Act as the technical authority for pre-sales and enterprise evaluations
- Demonstrate hardware and software capabilities at technical forums, conferences, and workshops
- Train customers on integration patterns, optimization workflows, and best practices
- Relay customer feedback to product and engineering teams
