





Metro location plus broad skills and mid-entry SRE title increase applicant density.
Core SRE skills are broadly transferable across industries, so background fit sensitivity is low.
Explicit years plus mandatory cloud, Kubernetes, IaC, and monitoring skills make filters moderately strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure stability, performance, and operational excellence of AI-driven healthcare cloud platforms by monitoring infrastructure, Kubernetes, CI/CD systems, and platform services.
Develop automation tooling and Infrastructure-as-Code to improve efficiency and reliability of cloud resources and deployment workflows.
Participate in incident response, troubleshoot production issues, and improve system scalability, resilience, and developer experience through collaboration with engineering teams.
Bachelor's or Master's degree in Computer Science, IT, Software Engineering, or related technical field.
1–3 years professional experience in Site Reliability Engineering, DevOps, Cloud Infrastructure Engineering, or related role.
Hands-on experience with Linux administration, scripting (Python, Bash, or Shell), cloud platforms (AWS, GCP, or Azure), containerization (Docker, Kubernetes), and Infrastructure-as-Code tools (Terraform or similar).
Familiarity with monitoring/observability tools (Prometheus, Grafana, ELK, Datadog), networking fundamentals, CI/CD pipelines, and version control (Git).
Experience operating in cloud-native, production environments with exposure to incident management and reliability engineering concepts (SLIs, SLOs, error budgets).
Practitioner of automation and operational excellence focusing on scalability, resilience, and improved developer productivity in distributed systems.
Comfortable collaborating across multiple engineering disciplines (Software, AI, Platform, Data Science) to drive continuous platform improvements and support.