Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, deploy, and operate scalable, resilient systems using serverless tech, event stream processing, and Kubernetes.
Develop and maintain monitoring, alerting, and incident response tools to ensure system reliability and uptime.
Automate operational tasks, troubleshoot complex distributed systems, and lead post-incident improvements.
Minimum Requirements
Minimum 5 years cumulative experience in Site Reliability Engineering, DevOps, Systems Engineering/Ops, or Software Development.
1-2 years experience in Bash scripting and systems administration or networking.
Demonstrated experience with Linux systems, containerization tools (Docker, Podman, Kubernetes), and managing large-scale distributed systems.
Bachelor’s degree preferred (Software Engineering, Computer Science, or related); equivalent experience and certifications considered.
Ideal Candidate Profile
Experienced in building robust, maintainable, and scalable cloud-native systems with ownership mindset.
Comfortable proactive identification and resolution of reliability issues independently with cross-team collaboration.
Familiarity with cloud platforms like Google Cloud Platform and Azure is highly desirable.
