





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Strong employer brand plus metro location and mid-level experience increases applicant competition density.
Role demands specialized LLM, GPU, and RAG experience, so candidates from other domains transfer poorly.
Explicit 5+ years, LLM training, GPU cluster and deployment expertise make hiring filters stringent.
Architect and deliver end-to-end generative AI solutions focused on Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) workflows.
Lead training, optimization, and scalable implementation of LLMs using NVIDIA hardware and software platforms.
Collaborate with customers and internal teams to understand requirements, provide technical leadership, and support pre-sales activities including workshops and demonstrations.
Master's or Ph.D. in Computer Science, Artificial Intelligence, or equivalent experience.
5+ years of hands-on experience in generative AI with strong emphasis on training and deploying Large Language Models.
Proven experience deploying and optimizing LLMs for production inference environments using frameworks like TensorFlow, PyTorch, or Hugging Face Transformers.
Strong knowledge of GPU cluster architecture and experience leveraging GPUs for accelerated LLM training and inference.
Experienced in architecting and delivering complex, scalable AI solutions specifically involving LLMs and RAG workflows.
Able to engage effectively with customers and cross-functional teams to translate business challenges into technical solutions and lead technical workshops.
Demonstrated proficiency with GPU-based AI workloads, model optimization, and deployment across cloud and on-premises environments.