Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Operate, monitor, and maintain Kubernetes-based cloud-native telecom platforms ensuring platform reliability and availability.
Manage and troubleshoot Kubernetes clusters, Istio service mesh, and associated networking and storage components.
Handle incident management including response, root cause analysis, and service recovery; create monitoring dashboards and maintain SLO/SLA reporting.
Minimum Requirements
Bachelor's degree in Computer Science, IT, Engineering, or equivalent.
4+ years of experience in Site Reliability Engineering, Cloud Operations, or Platform Operations.
Mandatory skills: Kubernetes administration and troubleshooting, Linux system administration, Istio service mesh operations, Prometheus, Grafana, AlertManager, Kubernetes networking, RBAC, Ansible, and storage administration.
Work Experience Required: 4+ years in relevant domains.
Ideal Candidate Profile
Experienced in managing production-critical, cloud-native infrastructure with strong operational ownership and incident response accountability.
Proficient in end-to-end Kubernetes cluster lifecycle management including deployment, configuration, networking, and storage.
Skilled in integrating monitoring and observability tools with a focus on operational excellence and scalability in telecom or similar environments.
