Site Reliability Engineer
Charger Logistics IncMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Ensure high reliability, scalability, and availability of a large-scale production platform supporting 24/7 logistics operations.
Define and monitor SLIs/SLOs and maintain observability through tools like Prometheus, Grafana, Loki, and Jaeger.
Manage Kubernetes clusters, automate operational tasks with scripting, maintain infrastructure as code, support CI/CD pipelines, and lead incident response including on-call duties.
Minimum Requirements
6–8 years experience in SRE, DevOps, or production infrastructure roles.
Strong hands-on experience with Kubernetes and Docker in production environments.
Proficiency in one major cloud platform (AWS, Azure, or GCP) and infrastructure as code tools like Terraform or Helm.
Experience with monitoring/observability tools (Prometheus, Grafana, Loki, Jaeger), scripting/coding in Python, Go or Bash, Linux and networking fundamentals, and CI/CD pipelines.
Ideal Candidate Profile
Experienced in managing distributed microservices platforms and production incident response with root-cause analysis.
Skilled in cloud-native infrastructure automation and scalable platform operations in 24/7 environments.
Comfortable working collaboratively with cross-functional teams (development, DevOps, QA) and handling on-call rotation responsibilities.
