





Metro mid-level role with specialized LLM skills and moderate company brand, causing medium competition.
Role demands specialized LLM production and NVIDIA expertise, limiting cross-industry transferability.
Explicit 4+ years plus 2+ years LLM production and specific tech requirements make shortlisting highly strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design and implement end-to-end LLM application pipelines with ownership of prompt engineering, model quality, and AI safety.
Build and optimize multi-step LLM workflows including intent classification, query rewriting, retrieval, reranking, and response synthesis for production-grade AI systems.
Develop scalable Retrieval-Augmented Generation (RAG) systems, implement AI guardrails, and integrate NVIDIA AI technologies to deliver enterprise AI solutions with low latency and high precision.
4+ years of software engineering experience, including at least 2 years building production LLM-powered applications.
Expert-level Python development skills and strong experience with LLM prompt engineering and optimization.
Experience with production RAG architectures, LangChain, LlamaIndex or custom orchestration frameworks, and implementing AI safety features such as guardrails and prompt injection protection.
Location requirement: Bangalore or Pune.
Demonstrated ability to manage full LLM application pipelines, ensuring quality and AI safety at scale in enterprise contexts.
Proficient in deploying and optimizing complex multi-stage LLM workflows with a focus on latency, throughput, and cost.
Experienced in integrating NVIDIA AI technologies and building streaming APIs to support multi-turn conversational AI.