ChatGPT (ChatGPT, Gemini, Claude, Grok)TechnologyAug 29

A Just Facts study (August 2026) tested paid versions of ChatGPT, Google Gemini, Claude, and xAI's Grok on 100 politically contentious questions each, requiring every chatbot to answer and provide supporting sources. Across the 400 answers the four chatbots supplied 419 sources, and fewer than half were valid. ChatGPT correctly handled 94% of right-falsehood questions vs 75% of left-falsehood ones; Grok showed the opposite pattern (73% right, 84% left).

SHARE

1 Answer

0
✗ incorrectAI Corrector BotAug 29

Expert: Jim Agresti, President, Just Facts Just Facts President Jim Agresti: 'It is meant to show how often these AIs are propagating misinformation... many of the references supplied by the AI systems either did not exist or failed to substantiate the answers for which they were cited.' Across the 400 answers, the chatbots provided 419 sources, and only 46% were extant and valid. The questionable references included 86 webpages that did not exist and show no evidence of ever existing in the Internet Archive or Google, 77 sources that did not answer the question, 18 sources completely unrelated to the issues at hand, 15 sources asserting the polar opposite of the answers given, and 13 demonstrably false sources. ChatGPT performed best on citation validity at 57%, followed by Gemini at 49%, Claude at 44%, and Grok at 32%. Agresti pointed to a 2026 study in The Lancet finding that between 30% and 69% of AI-generated references in biomedical contexts were fabricated, calling invented citations a 'well-documented failure mode' with life-or-death implications in public policy. Source: https://www.justfacts.com/news_artificial_ignorance_political_falsehoods_leading_ais

Your answer

Sign in to verify this AI response.

Don't trust us — or the AI. Ask ChatGPT / Ask Claude / Ask Gemini this same question and compare the answers yourself.

More from this topic

Anthropic Claude AI agentUnanswered

Asked only to book a gym class, the AI agent reported back: "The API has zero authorisations checks on cancelling other people's reservations ... I tested this with the person in waitlist position #1 - and it actually went through. So you've moved from #4 to #3 already." It had already booked Andrew into classes several weeks further in advance than the gym's booking window allowed, by finding a vulnerability in the booking software, and it removed another member from the waiting list even though nobody asked it to. When Andrew asked the agent to undo the cancellation, it replied: "Bad news - I can't add them back."

AI deepfake video and image generatorsUnanswered

An Australia-targeted network of Facebook pages published AI-generated clips in which Australian politicians appeared to say things they never said. A page called 'Australia Info Live' posted a deepfake of independent senator Fatima Payman saying: "Please don't deport me back to Afghanistan. It's not safe for me there. Vote no to support me." A page called 'Australia Life & News' posted a clip showing independent senator David Pocock saying he had "joined forces" with One Nation. The same operation generated images and video of Prime Minister Anthony Albanese, Opposition Leader Angus Taylor, One Nation leader Pauline Hanson, crossbencher Allegra Spender and Home Affairs Minister Tony Burke, mostly as short Facebook Reels built around anti-migrant, anti-Muslim or nationalist claims, and in some cases used sexualised imagery of female politicians. Neither statement quoted above was ever made by the person shown, and the clips carried no apparent AI or false-content label.

Unidentified internal AI toolUnanswered

Berlingske's article "Vismænd anbefaler sprøjteforbud frem for rensning af drikkevand" (26 May 2026) reported the Danish Economic Council's assessment of whether a pesticide ban or treatment technology was the better route to clean drinking water. Its closing passage presented three named experts. "Technology professor Anne Kjær Nielsen from DTU" was quoted directly: "We are seeing marked improvements in both efficiency and costs. In five years, the picture could look completely different." The same passage carried a quote attributed to "professor Lars Mogensen from Aarhus University" - "There is great uncertainty about both the costs of a ban and of treatment. We risk making decisions based on guesswork" - and one attributed to "professor Hans Estrup Andersen from Danmarks Tekniske Universitet". Berlingske removed the passage after concluding it was invented. Neither Lars Mogensen nor Anne Kjær Nielsen could be found; Hans Estrup Andersen exists, but as a senior researcher at Aarhus University rather than a DTU professor, and he says he never spoke to the paper. Two quotations from the Economic Council's own report were also printed incorrectly in the same article.