





Mid-level cloud role, metro location, broad cloud/Kubernetes skills and popular title increase competition.
Cloud infra skills are transferable across industries but require specific tooling and on-call experience.
Explicit 2.5–5 years plus mandatory AWS/Terraform/Kubernetes/on-call experience tightens selection filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own end-to-end incident management including on-call rotation, triage, mitigation, escalation, and resolution in a cloud infrastructure environment.
Maintain and enhance monitoring systems across environments including dashboards, alerting, log pipelines, and distributed tracing, focusing on reducing alert noise.
Provision and manage AWS cloud infrastructure using Terraform; debug Kubernetes issues; monitor CI/CD pipeline health; contribute to prototyping an AI-powered monitoring tool.
Work Experience Required: 2.5 to 5 years.
Strong experience with AWS, Kubernetes, Terraform, and monitoring/observability tools like Prometheus, Grafana, APM (Datadog/New Relic).
Mandatory ability to write clear, structured Root Cause Analyses for incidents readable by technical and non-technical stakeholders.
Location: Chennai, Tamil Nadu, India (On-site implied but not explicitly mentioned).
Demonstrates ownership and accountability by proactively resolving monitoring gaps and ensuring incident documentation completeness.
Strong analytical skills with understanding of distributed systems behavior under load through logs, metrics, and traces correlation.
Comfortable working in ambiguous and experimental settings, contributing actively to building new AI-based monitoring capabilities.