





Tier-1 brand, mid-level AI infra role, and metro location increase applicant competition.
Specialized AI inference and model-repository automation require domain-specific experience, reducing cross-industry transferability.
Explicit 5+ years and mandatory CI/CD, infrastructure, and programming skills enforce strict shortlisting.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, implement, and maintain infrastructure for running AI workloads across multiple inference backends (e.g., Llama.cpp, Ollama, Pytorch, WinML, TRT-RTX).
Develop automated systems for deploying AI applications, including model repository management and synchronization.
Build and maintain CI/CD pipelines for automated build and deployment; collaborate with Local AI developers to improve debugging and performance.
5+ years experience with application development in C#, Java, or similar languages.
B.Tech or higher degree in Computer Science, IT, Software Engineering, or related field.
Experience with scripting languages such as Python, Perl, or PHP.
Familiarity with databases, SQL, source control systems (Git, Perforce), and CI/CD tools (Jenkins).
Experienced in building backend automation systems with database management.
Hands-on exposure to CI/CD pipelines, Kubernetes, Docker, and visualization tools like Grafana or Kibana.
Familiarity with AI inference frameworks such as Llama.cpp, Ollama, PyTorch for automation deployment.