Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessTier-1 brand, metro location, popular backend title, and broad skillset increase candidate competition.
Specialized AI inference platform, service-mesh, and GPU workload expertise reduces cross-industry transferability.
Explicit 8+ years plus strong mandatory distributed-systems, cloud, and platform tech make filtering stringent.
Job Description
Structured overview of role & requirementsAbout This Role
Architect and build AI Inference Gateway platform for mission-critical machine learning workloads at scale.
Drive design and development of high-performance platform services including request routing, traffic shaping, load balancing, caching, and multi-tenant governance.
Lead technical vision and standards across multiple teams to ensure scalable, secure, and reliable AI inference infrastructure on cloud and on-premise.
Minimum Requirements
8+ years professional software development experience.
5+ years designing and operating large-scale distributed systems in production.
Strong programming skills in Go, Java, or Python.
Experience with Kubernetes, cloud platforms (AWS, Azure, or GCP), and building highly available infrastructure services.
Ideal Candidate Profile
Experienced in building AI/ML inference or model serving platforms at scale, including GPU-accelerated workloads.
Deep expertise in Kubernetes ecosystem and cloud-native distributed systems design.
Proven ability to lead architecture reviews, optimize system performance for large-scale networked services, and mentor senior engineers.
