AI/ML Site Reliability Engineer
LSEG (London Stock Exchange Group)Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessTier-1 brand, popular SRE role, and metro office increase candidate competition.
SRE skills are transferable but regulated finance and agentic AI platform experience raises domain specificity.
Mandatory Python, AWS, observability, SQL and regulated-delivery experience increases shortlisting strictness.
Job Description
Structured overview of role & requirementsAbout This Role
Build, operate, and ensure reliability of the Internal MCP Gateway and associated AI platform services, maintaining availability and performance targets.
Implement observability, monitoring, and audit capabilities for AI agent tool-calling and MCP traffic.
Collaborate with ML and Quality Engineers for testing and safe rollout, and contribute to platform architecture for federated MCP scale.
Minimum Requirements
Strong Python programming experience for platform services and tooling.
Experience with observability and monitoring tools (metrics, tracing, logging, dashboards).
Hands-on experience operating services on AWS cloud platform.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced Site Reliability Engineer with background in cloud-native reliability patterns and service lifecycle management in regulated or compliance-audited environments.
Comfortable working on AI platform components supporting enterprise-scale agentic AI services.
Skilled in operational analysis and SQL for audit queries and reporting within production-grade services.
