AI Chatbots Get Financial Questions Wrong More Than Half the Time, Study Finds
4 Articles
4 Articles
The models made calculation errors, among other things, and ignored upcoming tax rules, a new study has found.
Financial Times: AI chatbots give wrong answers to financial queries ‘most of the time’
Financial Times: AI chatbots give wrong answers to financial queries ‘most of the time’. “The most popular AI models from ChatGPT, Claude, Copilot, Grok and Gemini provided wrong answers to financial queries 57 per cent of the time on average, according to research from technology firm Saturn. When asked more complex questions, such as those which involved more than one calculation, the AI models made mistakes in 88 per cent of cases on average.…
AI Chatbots Get Financial Questions Wrong More Than Half the Time, Study Finds
A Saturn study tested 18 AI models, including ChatGPT, Claude and Gemini, on 121 financial questions and found they were wrong 57% of the time overall, and up to 99% wrong on harder ones. A separate PensionBee survey found most users would act on that advice without checking it first, even as card companies race to put AI agents in charge of real spending decisions.
Coverage Details
Bias Distribution
- There is no tracked Bias information for the sources covering this story.
Factuality
To view factuality data please Upgrade to Premium








