





Tier-1 brand, mid-level SRE role and Bangalore location create high candidate competition.
Role requires specific SRE/cloud/Kubernetes expertise so candidates from other domains have moderate transferability.
Explicit 5+ years SRE experience, Kubernetes, observability, cloud, and scripting requirements make filtering strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Accountable for maintaining and improving reliability, availability, and performance of large-scale GeForce NOW cloud gaming production services.
Own incident triage, troubleshooting, resolution, and participate in on-call rotations to ensure prompt restoration of customer-facing services.
Develop automation tools, improve operational workflows, drive observability initiatives, and collaborate across teams to enhance service SLOs and operational efficiency.
Bachelor's degree in Computer Science, Computer Engineering, Information Technology, or related field (or equivalent experience).
5+ years experience in Site Reliability Engineering or similar role supporting mission-critical production services in live-site environments.
Strong expertise in Kubernetes, containerization, microservices, distributed systems, and public cloud platforms (AWS, Azure, GCP).
Proficient in automation scripting/programming (Python, Go, Bash) and experience with observability tools (Prometheus, Grafana, ELK/OpenSearch).
Experienced in supporting large-scale, customer-facing cloud and gaming services with deep Kubernetes operational knowledge.
Proven track record driving incident response, postmortems, and continuous operational excellence initiatives.
Skilled in developing automation and observability solutions that improve engineering productivity and service reliability.