





Multiple amplifiers: mid-level generalist SRE role, broad tooling requirements, and Bangalore metro location drive high competition.
SRE skills transfer across industries but require platform-specific cloud and tooling experience, giving medium sensitivity.
Explicit 4–8 years plus mandatory observability, Kubernetes, AWS, IaC, scripting, and incident management requirements increase strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own and maintain the health and observability of MiQ's enterprise platform Sigma, including designing and implementing monitoring, alerting, and synthetic checks.
Drive platform reliability by defining SLIs/SLOs, optimizing release pipelines, automating deployment analysis, and improving incident response with on-call readiness.
Enhance SRE maturity through performance engineering, automation, and applying AIOps capabilities, ultimately owning observability standards and evangelizing best practices across engineering teams.
4 to 8 years of experience in Site Reliability Engineering, DevOps, or platform/production engineering supporting customer-facing systems.
Hands-on expertise with observability tools such as Grafana, Prometheus, Datadog, and experience with log/trace aggregation technologies like Loki or OpenTelemetry.
Experience operating workloads on AWS and Kubernetes (EKS preferred), with scripting/automation skills in Python, Bash, or Go, and familiarity with infrastructure-as-code and CI/CD pipelines.
Proven knowledge of SRE principles: SLIs/SLOs, error budgets, alert tuning, synthetic monitoring, incident management, and on-call processes.
Experienced SRE professional able to independently spearhead observability initiatives for a large-scale enterprise platform and reduce mean time to detect and resolve issues.
Skilled in integrating and optimizing monitoring solutions, balancing alert noise reduction with actionable insights, and collaborating effectively with cross-functional engineering teams.
Proficient in applying performance engineering and automation techniques with a proactive, ownership-driven approach to reliability that aligns with progressive delivery and platform stability goals.