





Tier-1 employer, metro Bangalore posting with broad SRE skillset attracts many qualified applicants.
High—requires deep SRE, cloud, and distributed systems experience, limiting cross-industry transferability.
Multiple mandatory cloud, SRE, and tooling requirements plus explicit 6+ years make filters stringent.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP environments.
Operate and support distributed data platforms and infrastructure components including Kubernetes, Kafka, Flink, Storm, Spark, and multiple databases (Cassandra, Elasticsearch, Redis, Postgres, ArangoDB).
Own incident management lifecycle including 24x7 on-call support, root cause analysis, runbook development, automation, capacity planning, and driving SRE best practices.
6–10+ years experience in DevOps, Site Reliability Engineering, or cloud infrastructure roles.
Bachelor’s or Master’s degree in Computer Science, Information Systems, or related field.
Strong hands-on experience with AWS or GCP cloud platforms and container orchestration technologies like Docker and Kubernetes.
Must be willing to work onsite primarily from an HPE office location.
Experienced in managing and troubleshooting multi-cloud, microservices-based production environments with strong automation and scripting skills (Python, Go, Rust, Shell).
Proficient in maintaining CI/CD pipelines, observability tooling (Prometheus, CloudWatch, Stackdriver), and configuration management (Ansible, Terraform).
Able to collaborate effectively with software engineering teams to resolve complex production issues and continuously improve operational processes.