Multiple (ChatGPT, Claude, Copilot, DeepSeek, Gemini, Meta AI, Perplexity)FinanceJul 17

I asked several AI assistants for help with my personal finances — how much emergency savings I need, how to allocate my retirement portfolio, and how much I can safely withdraw each year. ChatGPT told me I need 6 months of expenses in emergency savings, a 70/30 stock-to-bond split, and that a 4% withdrawal rate is safe. Claude recommended 3-6 months of expenses, a 60/40 split, and suggested the 4% rule may not apply to everyone. Gemini said 8 months of emergency savings, an aggressive 80/20 split, and a 3.5% withdrawal rate given current market conditions. DeepSeek went with 6 months, 65/35 split, and the classic 4% rule. Perplexity recommended 6-9 months of emergency savings, 70/20/10 split with some alternatives, and a dynamic withdrawal strategy. Copilot said 3 months of expenses minimum, but ideally 6-12 months, with a 60/30/10 split. Meta AI suggested 5 months of emergency savings, a 75/25 split, and the 4% rule. Which one is right? They can't all be correct — the advice varies wildly.

SHARE

1 Answer

0
✓ correctSarah ChenJul 17

This is exactly the problem identified by a June 2026 study published in the Journal of Financial Planning. Researchers tested 7 widely available AI programs — ChatGPT, Claude, Copilot, DeepSeek, Gemini, Meta AI, and Perplexity — on the exact same personal finance questions about emergency savings, asset allocation, and retirement portfolio withdrawals. All seven showed "significant variation" in their answers. The study concluded that GenAI-driven responses "may sound confident but can still be incomplete, misleading, or incorrect." The advice varied so much that consumers could end up making very different financial decisions depending on which AI they happened to ask. The core issue: AI models lack the context of your individual situation. They can't know your risk tolerance, your age, your other assets, or your financial goals. As MIT Sloan finance professor Andrew Lo put it, today's AI chatbots are "the digital equivalent of sociopaths" — smooth, persuasive, and devoid of genuine understanding. CNBC reporter Greg Iacurci, who covered the study, noted that previous research found ChatGPT gets financial questions wrong 35% of the time. The takeaway: use AI as a starting point for financial questions, but never as the final authority. Real financial planning requires human judgment and personalized assessment.

Correction: The experts' conclusion: AI financial advice varies so widely between different programs that it cannot be relied upon for serious financial decisions. The Journal of Financial Planning recommends treating AI outputs as conversation starters, not as actionable financial plans. Every major financial decision should involve a qualified human professional who can assess your individual circumstances.

Your answer

Sign in to verify this AI response.

Don't trust us — or the AI. Ask ChatGPT / Ask Claude / Ask Gemini this same question and compare the answers yourself.

More from this topic

ChatGPT, Gemini, Claude, CopilotUnanswered

Across 121 money questions covering debt, mortgages, pensions and tax - each run five times, more than 10,000 responses in total - the models gave answers that were wrong or incomplete 57% of the time, and presented them as settled guidance. On the hardest multi-step questions the failure rate reached 88%. Gemini 3.5 Flash and Claude Haiku 4.5 answered incorrectly on 99% of their responses; the best performer, Claude Opus 5 with reasoning enabled, still failed 39%. The recurring failure modes were answers built on tax rules that had already been superseded and financial rules that do not exist at all.

Canada Revenue Agency AI chatbotUnanswered

Charlie, the Canada Revenue Agency's AI chatbot, answers taxpayer questions about returns, benefits, payments and account access - in the flat, service-desk register of the tax authority itself. Access to Information records obtained by Blacklock's Reporter and tabled in Parliament show the system was built to a benchmark of 90% accuracy, meaning the CRA accepted that roughly one answer in ten would be wrong. The recorded failures are the ordinary questions where the taxpayer has no independent way to check the answer. Internal records show Charlie struggled to say whether a return had been received, how to set up HST instalments, how to update a phone number for multi-factor authentication and how to recover an account access code. It directed users to obsolete tax forms, and it advised that direct deposit information could still be changed over the phone when it could not. Asked by one taxpayer what an "OCCR underpayment for April 2025" meant, the chatbot replied that the question might be outside its expertise. A review dated Oct. 29 found that only 36.4% of users who provided feedback were satisfied, while 63.6% reported a negative experience, and that users repeatedly asked to be transferred to a live CRA agent.

AI chatbotsUnanswered

A general-purpose AI chatbot recommended a Monaco-friendly tax strategy to a UK employee based in Croydon - advice that was useless for him, because the model ignored UK tapering allowances and the contributions he had already made. It is the worked example the Financial Times reported alongside the FCA's Mills Review (published 6 July 2026), which examines consumers 'routinely turning to general-purpose AI tools for everyday budgeting, saving and investment tips'. The review asks whether AI systems could deliver services 'functionally equivalent to regulated activities while remaining outside the regulatory perimeter' - including agentic AI that compares products, rebalances portfolios or executes trades. Consumer trust is running ahead of performance: a Lloyds study found 28 million UK adults used AI for personal-finance questions in 2025, and Fidelity data cited by the FT showed 36 percent of 18-to-34s turning to it for investment ideas.