- Location
- Bengaluru, KA,IN, IN
- Type
- Full-time
- Department
- Engineering
- Seniority
- Senior
- Source
- Eightfold
Description
Architect end-to-end generative AI solutions with a focus on LLMs training , deployment and RAG workflows. Collaborate closely with customers to understand their language-related business challenges and design tailored solutions. Collaborate with sales and business development teams to support pre-sales activities, including technical presentations and demonstrations of LLM and RAG capabilities. Work closely with NVIDIA engineering teams to provide feedback and contribute to the evolution of generative AI software. Implement strategies for efficient and effective training of LLMs to achieve optimal performance. Design and implement RAG-based workflows to enhance content generation and information retrieval. Work closely with customers to integrate RAG workflows into their applications and systems and stay abreast of the latest developments in language models and generative AI technologies. Provide technical leadership and guidance on best practices for training LLMs and implementing RAG-based solutions. Expertise in training and fine-tuning LLMs using popular frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers. Experience leading workshops, training sessions, and presenting technical solutions to diverse audiences. Experience in deploying LLM models in cloud environments (e.g., AWS, Azure, GCP) and on-premises infrastructure.