





Tier-1 brand, remote posting, metro locations, and broad SRE skillset increase candidate competition.
Highly specific SRE/platform tools and enterprise streaming expertise limit cross-industry transferability.
Explicit 6+ years requirement and extensive mandatory tech stack make filters highly selective.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own platform reliability including availability, resilience, latency, and operational efficiency for event-driven cloud ecosystems.
Lead DevOps automation initiatives, CI/CD pipeline reliability (GitHub Actions), and observability stack management (Prometheus, Grafana, AlertManager, logging).
Drive cloud infrastructure maintenance, governance, capacity planning, incident leadership, and post-incident reliability improvements.
6+ years experience in SRE, platform engineering, DevOps, or advanced production support.
Strong hands-on expertise with Kubernetes (especially Azure AKS), CI/CD with GitHub Actions, and observability tools (Prometheus, Grafana, AlertManager).
Experience with streaming platforms including Confluent Kafka, Confluent Cloud, Azure Event Hub, AWS-MSK, and Apache Flink.
Work location onsite at Hyderabad or Bangalore (AT&T designated locations).
Senior to Lead-level individual contributor with 10 to 17 years experience in high-availability platform reliability roles.
Strong operational background with cloud-native platforms and event-driven streaming ecosystems, capable of leading high-severity incident response.
Skilled in automation (Python scripting), platform governance, end-to-end production troubleshooting of complex distributed microservices and streaming infrastructure.