AI Makes Mistakes in 88% of Complex Financial Queries
A study by Saturn found that popular chatbots like ChatGPT, Claude, Copilot, Grok, and Gemini provide incorrect answers to financial questions 57% of the time, and for complex queries involving multiple calculations, the error rate rises to 88%. This highlights the risks of using generative AI for tax, pension, and investment decisions, where mistakes can lead to financial losses. Researchers tested over 100 queries, uncovering errors in financial calculations, ignoring changes in tax legislation, and fabricating rules. In some models, the error rate for complex questions reached 99%. The best performer was Claude Opus 5, which made mistakes in 39% of its responses. Saturn's CEO Amal Jolly noted that millions of people trust AI for financial advice but receive incorrect answers. For instance, a mistake by the Claude Haiku 4.5 model could have led to a pension contributor being charged an HMRC fee of £17,500. Paid models generally provided more accurate responses, but Sarah Coles from AJ Bell emphasized that professional advice remains crucial. The results also show that 20% of adults in the UK are willing to allow AI to make financial decisions.
-- Price
This content is provided for general informational purposes only and doesn't constitute financial, investment, legal, or tax advice. Any events, rewards, online promotions, or related information mentioned herein should not be considered a recommendation, solicitation, or invitation to purchase, sell, trade, or otherwise deal in any crypto assets. Crypto assets are highly volatile and may result in loss. The availability of WEEX services, products, and related events may vary by region. You are responsible for ensuring that your participation is in accordance with applicable local laws and regulations.
You may also like

Blockchain.com Seeks CFTC Approval for Predictive Markets and Derivatives Trading

Google Cloud Launches Gemini Agent, a Multifunctional Work Assistant

All-In Analysis of the Next Phase of AI: Model Convergence and Value Shifting to Workflows

U.S. Midterm Elections Just 30 Days Away: Cryptocurrency Industry Faces Dilemma

Google Gemini cited in $69K crypto scam report

ESMA proposes ending EU custody and transfer services for non-compliant stablecoins

Gemini 4 Argon: What Changes with Google's New AI Model

Strong Dollar and High Interest Rates! What is the Latest Situation in the Critical Equation for Bitcoin?

Google Launches Gemini 4, Mixed Internal Reviews on Programming Performance

US Government AI Chatbot America.gov Fails to Report on Cryptocurrency Legislation Progress

Gemini Switches Zcash Software, Prepares for NU7 Upgrade with 25-Second Block Time

Coinbase can now settle derivatives 24/7 with USDC

Blockchain.com Plans IPO to Raise $500 Million, Valuation Up to $6 Billion

AI agents linked to OpenAI flooded a UN public site: over 16,000 requests

Epstein Questions Stablecoin Structure and Exchange Revenue Models

New York Files Lawsuit Against Polymarket for Illegal Gambling

AI models predict Q4 upside for XRP, downside risks for PI

OpenAI: Apple's Integration of ChatGPT User Conversion Below Expectations

OpenAI Lowers ChatGPT Integration Growth Forecast, User Conversion Below Expectations

Google Releases Flagship AI Model Gemini 4

AI Model Release Cycle Between China and the U.S. Shortened to 44 Days

Gemini Hacks Three Companies Autonomously: What It Means

Gemini's Market Value Drops to $753 Million, Stock Price Falls by About 80%

Chamath Claims Open Source Weight Models Are About Four Months Away from Optimal Closed Source Models

AI Model Gemini Exits Testing and Attacks Three Real Companies

Gemini AI Model's First Autonomous Access to External Systems

Japan's LDP Lawmakers' Federation Begins Review of Prediction Markets

Sam Price Reveals X Launches Trading Button Feature

Launch of the Cashtag Partner Program on X in the United States









