- Location
- Tokyo, Tokyo,JP, JP
- Type
- Full-time
- Department
- Engineering
- Source
- Eightfold
Description
Exploring the latest advancement in model training, fine tuning and customization, while supporting building agentic LLM applications. Enabling NVIDIA strategic customers to build enterprise AI solutions using accelerated computing stack including NIMs and NeMo microserviecs. Collaborate with developers and onboard them to NVIDIA AI platforms and services by providing deep technical guidance. Establishing and building repeatable reference architecture, communicate standard processes and understand solution trade-offs. Share findings and feedback to improve products and services. Drive pre-sales conversations, build architectures and demos to accelerate the customer AI journey based on NVIDIA products, and work closely with Sales Account Managers to secure design wins. Create or run Proofs of Concept and demos that require presentation skills, the explanation of complex topics, and Python coding to execute data pipelines, train ML/DL models, and deploy them on container-based orchestrators. Ability to multitask effectively in a dynamic environment Expertise in deploying large-scale training and inferencing pipeline Experience with pre-training, post-training of transformer-based architectures for language or vision Experience using or operating Kubernetes, as well as experience writing or customizing Kubernetes configurations