Senior Production Support Engineer(4+ Years in Azure and AWS infrastructure support and SQL Server administration)
FISMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessMid-level, popular SRE/devops role with broad cloud, Kubernetes, and SQL requirements attracts many candidates.
Cloud, Kubernetes, SQL Server, and SRE skills transfer easily across industries.
Explicit 4+ years and mandatory cloud, Kubernetes, and SQL Server expertise create strict screening.
Job Description
Structured overview of role & requirementsAbout This Role
Provide L2/L3 production support and lead troubleshooting across applications, databases, cloud infrastructure (Azure, AWS), networking, and Kubernetes.
Participate in incident response, conduct Root Cause Analysis (RCA), develop corrective actions and maintain operational runbooks and recovery procedures.
Improve service reliability and operational efficiency by enhancing monitoring/observability (Dynatrace, Azure Monitor, AWS CloudWatch), automating tasks, and partnering with Engineering, Infrastructure, Security, and Product teams.
Minimum Requirements
Bachelor’s degree in Computer Science, IT, Engineering or equivalent practical experience.
Proven experience in Production Support, Site Reliability Engineering, or similar operational support role with strong focus on Azure and AWS infrastructure support.
Hands-on experience with Microsoft SQL Server administration including troubleshooting and performance tuning.
Experience supporting Kubernetes (AKS), GitHub, ArgoCD, and strong understanding of networking concepts (DNS, TCP/IP, routing, firewalls). Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced with cloud-native and microservices architectures, and skilled in multi-layer troubleshooting across infrastructure, applications, and databases.
Comfortable applying SRE principles including observability, automation, error budgets, SLOs, blameless post-mortems, and operational excellence.
Able to drive operational improvements by collaborating with development and DevOps teams, and proactive in enhancing platform stability and automation.
