Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own end-to-end architecture, deployment, and scaling of on-prem AI/ML platforms including data pipelines and production monitoring.
Design and manage CI/CD pipelines, containerization (Docker/Kubernetes/OpenShift), and infrastructure for ML/LLM models on GPU-enabled systems.
Ensure production reliability, observability (Prometheus, Grafana, ELK), and lead technical mentorship and cross-team collaboration.
Minimum Requirements
6+ years experience in MLOps, DevOps, or Platform Engineering.
Strong skills in Python, Bash scripting, Linux (preferably RHEL), Docker, Kubernetes/OpenShift, and CI/CD tools (Jenkins/GitLab CI).
Experience deploying ML/LLM models on on-prem GPU infrastructure with real-time and batch data integration using Kafka.
Must understand and explain AI/ML platform architecture, OpenShift AI ecosystem, and container orchestration; knowledge of SQL and data pipelines mandatory.
Ideal Candidate Profile
Senior-level hands-on platform architect comfortable managing both ML systems and underlying infrastructure.
Experienced in scaling production ML systems in restricted or air-gapped enterprise environments.
Able to clearly explain complex integration patterns (Kafka, Gunicorn) and drive automation and best practices for scalable, secure ML platforms.
