Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, implement, and maintain highly available, scalable, and resilient infrastructure and DevOps systems including CI/CD pipelines and automation workflows.
Establish and maintain service reliability engineering (SRE) practices such as monitoring, alerting, incident response, and defining service-level objectives (SLOs).
Apply AI/ML techniques to operational workflows (AIOps) to improve anomaly detection, alert correlation, root-cause analysis, and automate incident management.
Minimum Requirements
Bachelor's degree in Computer Science, Engineering, Information Systems, or related field, or equivalent experience.
Hands-on experience with cloud platforms (Microsoft Azure, AWS, or Google Cloud Platform).
Experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or Cloud Operations roles with CI/CD pipeline design and deployment automation.
Hands-on experience using LLMs, AI models, GitHub CoPilot, and LLM with Context Engine.
Ideal Candidate Profile
Demonstrated expertise in SRE and DevOps practices, including service reliability, observability, and automation in cloud-native environments.
Experience applying AI/ML to IT operations for automation and predictive analytics with a strong focus on enterprise-scale tooling and operational excellence.
Proficient in scripting/programming languages such as Python, Bash, JavaScript, Go, or Java, and knowledgeable in containerization and orchestration technologies (Docker, Kubernetes).
