





Tier-1 brand, mid-level (5+ years) role, and metro location increase applicant density.
Role requires niche system programming and AI-inference infrastructure skills, limiting cross-industry transferability.
Explicit 5+ years requirement plus mandatory systems programming and infrastructure experience raises strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, implement, and maintain system software and infrastructure to run AI workloads across various inference backends like Llama.cpp, Ollama, PyTorch, vLLM, WinML, TRT-RTX.
Develop and maintain infrastructure for automated build, integration, deployment, and qualification of AI models and applications, including local model repository management and synchronization.
Analyze large datasets and system metrics to build data processing and visualization tools, and collaborate on implementing system-level debugging, diagnostics, and performance optimization solutions.
5+ years software development experience with strong programming skills in C/C++, C#, Java, or equivalent systems programming languages.
B.Tech or higher degree in Computer Science, IT, Software Engineering, or a related field.
Experience with system APIs, multithreaded software, process and resource management, debugging, performance analysis, databases/SQL, source control (Git, Perforce), and CI/CD infrastructure (Jenkins).
Outstanding written and verbal communication skills; Work Experience Required: 5+ years
Experienced in building robust system software, runtime infrastructure, or scalable backend components for AI or similar workloads.
Skilled in system profiling, performance optimization, debugging, telemetry and observability tools (e.g., Nsight Systems, Grafana, Kibana).
Hands-on expertise with Linux/Windows system programming, containerization (Kubernetes, Docker), CI/CD pipelines, and integrating/debugging AI inference frameworks (Llama.cpp, PyTorch, TRT-RTX).