Senior LLM Inference Performance and Evaluation Engineer

Bitdeer is a technology company providing Bitcoin mining solutions.

0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/25/2026

Singapore, SG
About Bitdeer

Bitdeer provides full-spectrum Bitcoin mining and high-performance computing solutions, including SEALMINER mining equipment, Minerbase cooling containers, cloud mining, co-mining, mining management applications, mining rights marketplaces, and large-scale data center operations. The company also offers AI cloud infrastructure with GPU computing, model training and deployment capabilities, and turnkey AI data center solutions for enterprise customers and developers. Bitdeer is headquartered in Singapore and operates globally.

View jobs by Bitdeer

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You build benchmark pipelines for LLM inference performance, create model launch gates, maintain representative workloads, compare model and runtime options, automate regression detection, and collaborate with runtime engineers and SRE to identify bottlenecks and define service objectives and alert thresholds.

Requirements

  • 5+ years of experience in ML infrastructure, performance engineering, model evaluation, QA automation, or backend testing for production systems.
  • Experience with TTFT, TPOT/ITL, request latency, token throughput, concurrency, and GPU-utilization metrics.
  • Strong Python skills.
  • Go experience preferred.
  • Familiarity with OpenAI and Anthropic APIs, vLLM, Dynamo, SGLang, Triton-style servers, and Kubernetes test environments.
  • Ability to design statistically meaningful tests and communicate tradeoffs.
  • Experience building dashboards, reports, and release gates.

Responsibilities

  • Build benchmark pipelines for latency, throughput, concurrency, error rate, and GPU utilization.
  • Create model launch gates for API compatibility, streaming, tool calling, reasoning, multimodal behavior, and long-context cases.
  • Maintain synthetic, replayed, and customer-like workloads.
  • Compare model, runtime, and provider options and recommend routing, fallback, pricing, and capacity decisions.
  • Automate regression detection in CI/CD and staging.
  • Identify bottlenecks and verify performance improvements.
  • Convert benchmark results into SLOs and alert thresholds.
  • Build dashboards, reports, and release gates.

Benefits

  • Inclusive and respectful work environment.
  • Opportunity to contribute directly to the future of the digital asset industry.
  • Involvement in new projects and process development.
  • Personal accountability, autonomy, fast growth, and learning opportunities.
  • Training and mentoring opportunities.
  • Attractive welfare benefits.