Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Deliver 24x7x365 tier two support and escalation management for AWS and Azure cloud environments with accountability for maintaining system stability and operational excellence.
Implement and maintain reliability engineering practices including SLIs/SLOs, advanced troubleshooting, incident management, and root cause analysis across multi-cloud infrastructure.
Automate complex operational tasks using infrastructure as code, CI/CD pipelines, and scripting; manage containerized workloads and observability solutions to improve operational efficiency across AWS and Azure.
Minimum Requirements
5+ years of hands-on experience in DevOps, SRE, or production operations roles.
Deep expertise in either AWS (EC2, RDS, S3, IAM, VPC, CloudWatch) or Azure (VMs, Azure SQL, Storage, Azure AD, Virtual Networks) cloud platform; working knowledge of the other cloud platform is expected.
Proficiency in infrastructure as code tools (Terraform preferred), CI/CD pipelines, scripting languages (Python, PowerShell, Bash), containerization (Docker, Kubernetes), and cloud-native monitoring/logging tools.
Location requirement: Mumbai; Education: Any Graduate.
Ideal Candidate Profile
Experienced operating production cloud environments with proven ability to handle complex cloud escalations and multi-cloud incident response.
Strong operational ownership mindset with demonstrated skills in automating operational workflows and driving reliability best practices at scale across AWS and Azure multidomain environments.
Comfortable working in 24x7 managed services setup supporting customer-facing infrastructure with collaborative communication between engineering and client teams.
