Senior LLM Inference Performance and Evaluation Engineer
Bitdeer Technologies Group · Singapore, SG / Penang, MY / Taiwan ·
- Seniority
- Senior
- Category
- AI engineering
- Experience
- 5+ years
Bitdeer Technologies Group · Singapore, SG / Penang, MY / Taiwan ·
bitdeer · Singapore, Singapore
NineTwoThree AI Studio · Spain
horizon3ai · US, Remote
NVIDIA · US, CA, Santa Clara
You build benchmark pipelines for LLM inference performance, create model launch gates, maintain representative workloads, compare model and runtime options, automate regression detection, and collaborate with runtime engineers and SRE to identify bottlenecks and define service objectives and alert thresholds.
Nebius · United States