- Location
- Shanghai, Shanghai,CN, CN · Beijing, Beijing,CN, CN
- Type
- Full-time
- Department
- Engineering
- Source
- Eightfold
Description
Collaborate with the deep learning community to implement the latest algorithms for public release in Tensor-RT. Develop highly optimized deep learning kernels for inference Do performance optimization, analysis, and tuning Work with cross-collaborative teams across automotive, image understanding, and speech understanding to develop innovative solutions Occasionally travel to conferences and customers for technical consultation and training Performance modelling, profiling, debug, and code optimization or architectural knowledge of CPU and GPU