Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and deploy production-ready Generative AI and LLM-powered applications including RAG systems, chatbots, AI assistants, and automation tools.
Develop and optimize scalable AI backend services and APIs using Python and frameworks like FastAPI or Flask, integrating with LLM platforms such as OpenAI and Anthropic.
Monitor, troubleshoot, and improve AI system performance, reliability, latency, and cost while collaborating with engineering, product, and operations teams globally.
Minimum Requirements
Minimum 3+ years of software engineering or AI/ML engineering experience.
Strong proficiency in Python and practical experience with LLMs, embeddings, vector databases, and RAG pipelines.
Experience developing backend services/APIs using FastAPI, Flask, or similar frameworks and integrating LLM APIs (OpenAI, Anthropic, Azure OpenAI).
Must work full Pacific Time schedule (7:00 AM–4:00 PM PT, approx. 7:30 PM–4:30 AM IST), remote India location, with flexibility for occasional weekend work.
Ideal Candidate Profile
Experienced in taking AI use cases from requirements through to production deployment, demonstrating measurable system reliability and optimization.
Deep hands-on experience with Generative AI toolkits including RAG, semantic search, AI prompt engineering, and vector search databases.
Comfortable working asynchronously and cross-functionally in a fully remote, globally distributed team adapting to evolving AI project priorities.
