





Tier-1 employer, popular mid-level SRE role and metro location increase candidate competition.
SRE and observability skills are transferable, though finance compliance adds moderate domain bias.
Explicit 5+ years and mandatory SRE, observability, Python and cloud skills make screening strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead development of scalable, resilient AI/ML data platform solutions within JPMorgan Chase's AI/ML Data Platforms team.
Coordinate and manage incident resolution including root cause analysis and production changes for site reliability.
Mentor and guide team members, driving adoption of AI-assisted engineering practices and strategic improvements in code quality and operational outcomes.
5+ years of applied software engineering experience with formal training or certification.
Proficiency in site reliability engineering principles, incident management, and observability tools like Grafana, Dynatrace, Prometheus, Datadog, Splunk.
Strong understanding and practical experience with SLI/SLO/SLA, error budgets, Python or PySpark for AI/ML modeling, and automation to reduce toil.
Hands-on experience in system design, resiliency, testing, operational stability, network topologies, and awareness of compliance/risk controls.
Experienced leader in site reliability engineering or production support with AWS Cloud, Databricks, Snowflake, or similar platforms.
Demonstrated ability to implement and drive AI-assisted software development tools and responsible AI usage within engineering teams.
Able to manage cross-functional collaboration and mentoring to embed strategic changes in large, complex, global technology environments.