Engineer II - Site Reliability (Hybrid, IND)
CrowdStrike, Inc.Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Operate and maintain the Temporal production workflow orchestration platform across multiple Kubernetes clusters and regions, ensuring high availability and performance.
Automate deployment, upgrades, scaling, and monitoring processes to reduce manual efforts and support capacity planning and performance tuning.
Collaborate with internal teams to onboard users, troubleshoot issues, contribute to runbooks, dashboards, and improve observability and incident response.
Minimum Requirements
3+ years experience in DevOps, SRE, platform engineering, or related infrastructure roles with hands-on production system operations.
Familiarity with Kubernetes fundamentals including deploying services, using kubectl, and debugging cluster issues.
Experience using Helm for application deployment and basic troubleshooting of Helm charts.
Some experience with infrastructure-as-code or GitOps tools (e.g., Terraform, Ansible, FluxCD, ArgoCD) and working knowledge of cloud platforms (AWS or GCP).
Ideal Candidate Profile
Early to mid-career platform engineer looking to deepen operational expertise and automation skills within distributed, stateful infrastructure.
Comfortable working with Kubernetes-based systems, Helm, scripting (Bash, Python, or Go), and have foundational knowledge of databases like PostgreSQL.
Able to collaborate with engineering teams to support and improve a critical workflow orchestration system and contribute to continuous improvement of operational processes.
