





High due to Tier-1 brand, popular SRE role, metro location, and broad skill requirements.
Medium because cloud-native SRE skills transfer across industries but require specific streaming and database expertise.
High due to explicit 6–10+ years requirement and many mandatory cloud, Kubernetes, and IaC skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP environments.
Operate and support infrastructure components and distributed data platforms including Kubernetes, Kafka, Flink, Storm, Spark, and multiple databases such as Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB.
Own incident management lifecycle including detection, mitigation, root cause analysis, post-incident reviews, and participate in 24x7 on-call rotation for multi-cloud production environments.
6–10+ years of experience in DevOps, Site Reliability Engineering, or cloud infrastructure roles.
Bachelor's or Master's degree in Computer Science, Information Systems, or related field.
Strong hands-on experience with cloud platforms AWS or GCP, containerization and orchestration (Docker, Kubernetes), and managing distributed systems (Kafka, Cassandra, Elasticsearch, Spark, Flink, Storm).
Onsite work requirement at HPE office.
Experienced in managing complex multi-cloud (AWS and GCP) infrastructure with emphasis on operational reliability and performance.
Proven ability to handle end-to-end incident management including 24x7 on-call support and post-incident analysis.
Skilled in automation, monitoring, and collaborating closely with software engineering teams to maintain distributed data platforms and microservices.