Site Reliability Engineer II - Python, Observability, AWS, Terraform
JPMorgan Chase & Co.Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Independently execute and eventually design small to medium projects focused on system reliability and operational stability.
Develop and maintain high-quality, maintainable code to solve business problems, reducing operational toil through automation and system improvements.
Implement and improve observability systems (SLI/SLO/SLA), utilize AI tools for incident triage and root cause analysis, and manage incident resolution with cross-functional teams.
Minimum Requirements
2+ years of applied experience with formal training or certification in software engineering concepts.
Proficient in site reliability engineering principles including SLI/SLO/SLA and error budgets, with practical AWS experience (EC2, S3, RDS, VPC, IAM, networking).
Proficient in Python or Java/Spring Boot programming, experience with CI/CD tools such as Jenkins, GitLab, or Terraform.
Experience with observability tools like Grafana, Dynatrace, Prometheus, Splunk and AI-assisted tools for operational workflows.
Ideal Candidate Profile
Experienced Site Reliability Engineer comfortable working independently on reliability projects and collaborating across teams.
Technically strong with demonstrated skills in AWS cloud environments, observability implementations, and incident management.
Practically skilled in leveraging AI-assisted tools to reduce operational toil and optimize monitoring and alerting systems.
