Machine Learning Quality Engineer
London Stock Exchange Group
Overview
In this role you contribute to the CE AI team by defining and applying rigorous quality and governance for enterprise AI tools. You will build automated test pipelines and evaluate agentic systems to ensure safety, correctness, and reliable performance. You collaborate with ML engineers to embed testability in MCP servers and Skills and produce evidence for compliance gates. This role enables scalable, governed AI capabilities across LSEG’s product teams and platforms.
Pay / Benefits- healthcare
- retirement planning
- paid volunteering days
- wellbeing initiatives
- Define and implement evaluation frameworks for MCP tools and Skills (correctness, safety, regression)
- Build and maintain automated test pipelines for agentic behaviours (tool invocation, multi-step workflows)
- Evaluate and mitigate agentic failure modes (hallucination, tool misuse, invalid inputs, latency)
- Produce testing evidence for Permit to Build (PTB) and Permit to Operate (PTO)
- Partner with ML Engineers to embed testability and evaluation hooks into MCP servers and Skills
- Help define the long-term quality and governance model for federated MCP contributions across LSEG
- Strong Python experience for test harnesses and automation
- Experience with LLM and RAG evaluation frameworks or custom evaluation pipelines
- Test automation expertise covering unit, integration, and regression testing
- Understanding of agentic system risks and failure modes
- Ability to assess solutions against governance, security, and audit expectations
- Experience working in regulated or highly governed engineering environments
- Collaborative mindset
- Problem-solving and analytical thinking
- Communication and cross-functional teamwork
- Python programming
- LLM/RAG evaluation frameworks or custom evaluation pipelines
- Test automation (unit/integration/regression)
Reference: WJ-747_30139039