Few banks formally evaluate GenAI human-in-the-loop controls
Benchmarking: G-Sibs and challengers use tools to test controls efficacy; others rely on judgment.
Only one in five banks use tools to formally monitor the effectiveness of the human-in-the-loop (HITL) controls they apply to generative AI or large language models (LLMs), Risk Benchmarking’s inaugural Model Risk Management study finds.
Four of the six global systemically important banks (G-Sibs) answering said they measure HITL effectiveness using key performance indicators to track metrics such as hit rates, escalations, and error detection, along with two challenger banks.
A further six banks
Only users who have a paid subscription or are part of a corporate subscription are able to print or copy content.
To access these options, along with all other subscription benefits, please contact info@waterstechnology.com or view our subscription options here: https://subscriptions.waterstechnology.com/subscribe
You are currently unable to print this content. Please contact info@waterstechnology.com to find out more.
You are currently unable to copy this content. Please contact info@waterstechnology.com to find out more.
Copyright Infopro Digital Limited. All rights reserved.
As outlined in our terms and conditions, https://www.infopro-digital.com/terms-and-conditions/subscriptions/ (point 2.4), printing is limited to a single copy.
If you would like to purchase additional rights please email info@waterstechnology.com
Copyright Infopro Digital Limited. All rights reserved.
You may share this content using our article tools. As outlined in our terms and conditions, https://www.infopro-digital.com/terms-and-conditions/subscriptions/ (clause 2.4), an Authorised User may only make one copy of the materials for their own personal use. You must also comply with the restrictions in clause 2.5.
If you would like to purchase additional rights please email info@waterstechnology.com
More on Emerging Technologies
Anna’s new digital token identifiers, TS Imagine offers prediction market data, and more
The Waters Cooler: A recap of the major tech and data news from the past week in the capital markets.
Waters Wavelength Ep. 356: When software fails
This week, Tony and Shen discuss a recent trade surveillance software failure.
A tidal wave of token costs threatens landfall
Budgeting for AI was never “easy,” but as financial firms rely more heavily on agents, soaring token usage is forcing them to rethink the economics of modern enterprise AI.
Can pairing autonomous AI agents with digital cash transform finance?
A new Moody’s paper highlights opportunities and risks associated with AI agents and digital cash.
Is this tokenization’s golden opportunity?
The Waters Wrap: More initiatives around tokenizing assets are coming to fruition. Nyela asks: Is the market ready?
LSEG caps free prompts as it moves to increase AI revenues
New AI tools were the main focus of the exchange group’s mid-year progress update.
EDMA begins rollout of AI certification framework
The EDM Association is expanding its CDMC framework to include analytics and AI.
LSEG plans new venue launch, Talos integrates with Kalshi, and more
The Waters Cooler: A recap of the major tech and data news from the past week in the capital markets.