Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Maintain and improve reliability, scalability, and performance of cloud infrastructure across AWS, Azure, and GCP.
Lead incident response including root cause analysis and postmortems to enhance uptime and resilience.
Design, deploy, and optimize cloud resources using Infrastructure as Code (Terraform, Ansible) and implement monitoring/observability frameworks.
Minimum Requirements
Proficient programming/scripting skills in Python, PowerShell, Bash, or equivalent for automation.
Hands-on experience with AWS, Azure, or GCP cloud platforms including VPCs, IAM, serverless, and Kubernetes.
Experience with Infrastructure as Code tools such as Terraform and Ansible.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Strong expertise in system reliability engineering with ability to handle incident management under pressure.
Experienced in multi-cloud environments and hybrid cloud deployments with knowledge of service-level objectives (SLOs) and error budgets.
Collaborative with ability to communicate technical concepts at all organizational levels and mentor junior engineers.
