Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own lifecycle management of firmware, kernel, OS, and software patching across managed hosting and application environments to reduce technical debt.
Design, build, and operate automation for provisioning, deployment, patching, remediation, and release pipelines with human-in-the-loop approval gates.
Define, instrument, and track SLIs/SLOs for platform reliability; provide frontline technical response and remediation during platform incidents.
Minimum Requirements
3–5+ years experience in platform engineering, systems engineering, SRE, or infrastructure operations.
Advanced Linux systems administration with kernel and firmware-level expertise.
Strong experience in Kubernetes, Docker, container orchestration, infrastructure-as-code tools (Terraform, Ansible, Puppet/Chef), and CI/CD pipelines.
Proven skills in SLI/SLO definition and tracking using observability tools (Grafana, Datadog, Prometheus) and experience as a technical responder in production incidents.
Ideal Candidate Profile
Experienced in enterprise-scale or customer-impacting production environments with hybrid or cloud-hosted infrastructure.
Technical leadership capability including mentoring and cross-team communication on platform architecture and operational readiness.
Strong automation and scripting skills (Python, Bash, Go) coupled with expertise in platform modernization and reliability engineering practices.
