- Location
- Taipei, Taipei City,TW, TW · Hsinchu, Hsinchu City,TW, TW
- Type
- Full-time
- Department
- Engineering
- Source
- Eightfold
Description
Craft and develop robust inference software that can be scaled to multiple platforms for functionality and performance Performance analysis, optimization, and tuning for Large Language Models (LLMs) Closely follow academic developments in the field of artificial intelligence and feature update TensorRT-LLM Provide feedback into the architecture and hardware design and development Collaborate across the company to guide the direction of deep learning inference, working with software, research and product teams Publish key results in scientific conferences Architectural knowledge of CPU and GPU Experience working with deep learning frameworks like PyTorch and HuggingFace