- Location
- Santa Clara, CA,US, US · Redmond, WA,US, US
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Source
- Eightfold
Description
As a member of our deep learning architecture team, you will contribute to features that help next-generation GPUs advance the state of AI. This position requires you to keep up with the latest DL research and collaborate with diverse teams (internal and external to NVIDIA), including DL researchers, hardware architects, and software engineers. Your day to day work will include analyzing the behavior of various deep learning methods, proposing new features to accelerate or enable various methods, and studying the benefits of the proposed features. Performance analysis and optimization; Experience with LLM workloads, including performance tuning considerations such as parallelization and fusion strategies;` Experience with core deep learning kernels such as matrix multiply, attention, and communication convolution Experience with GPU computing (CUDA) Experience with deep learning frameworks like PyTorch