Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessTier-1 employer, popular devops role in Hyderabad with broad skills increases competition.
Core automation and observability skills transfer well, though regulated life-sciences experience is preferred.
Explicit 10+ years, mandatory IaC/observability skills and regulated environment make filters strict.
Job Description
Structured overview of role & requirementsAbout This Role
Develop, maintain, and ship production-grade automation that executes validated IT remediation procedures end-to-end with built-in rollback and exception handling.
Translate SRE-authored remediation runbooks into codified, executable automation with full observability including metrics, logs, traces, SLIs/SLOs, dashboards, and alerts.
Collaborate with SRE and Operations teams to validate automation safety, ensure continuous improvement, and manage automation autonomy graduation criteria without causing critical incidents.
Minimum Requirements
Minimum 10+ years of hands-on automation engineering experience in enterprise IT operations building production automation or scripted remediation.
Strong proficiency with observability tools (e.g., Prometheus/Grafana/OpenTelemetry, Splunk, Datadog) including designing SLIs/SLOs and alert tuning.
Practical programming/scripting skills (Python, PowerShell, Bash) and experience with infrastructure as code/configuration management tools (Ansible, Terraform, or equivalent).
Bachelor's degree or higher in Computer Science, Information Technology, or closely related field; onsite based in Hyderabad with flexible shift availability.
Ideal Candidate Profile
Experienced in applying SRE principles such as error budgets, blameless postmortems, and graduated automation rollouts in regulated, audit-ready environments (life sciences preferred).
Demonstrates engineering rigor in building safe, observable, and well-documented automation that others can maintain and extend independently.
Proven ability to work cross-functionally with SRE, Reliability, and Operations teams to validate and evolve automation safely at scale with zero P1/P2 incident causation.
