Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Fine-tune, develop, and deploy large language models (LLMs) and NLP applications optimized for specific use cases.
Design, maintain, and ensure high performance and scalability of Python API backends serving AI/ML models in production.
Experiment and optimize model accuracy, efficiency, and latency through advanced training strategies and hyperparameter tuning while aligning AI capabilities with business goals.
Minimum Requirements
Minimum 4 years of overall experience with at least 1 project involving fine-tuning large language models.
Strong experience in building and fine-tuning domain-specific LLMs (e.g., GPT, LLaMA, Mistral, T5) and deploying them.
Proficiency in Python and AI/ML frameworks such as PyTorch or TensorFlow.
Experience with Retrieval-Augmented Generation solutions using vector databases (Pinecone, ChromaDB) and Python API frameworks (FastAPI, Flask, Django).
Onsite work required in Bangalore or Chennai (Hybrid Role).
Ideal Candidate Profile
Experienced in cutting-edge generative AI and LLM fine-tuning with demonstrated deployment in production environments.
Skilled in building scalable AI-powered APIs with a strong emphasis on model optimization and reliability.
Comfortable working in startup-like, fast-paced environments building state-of-the-art technology with product-aligned focus.

