Multiple (Gemini, ChatGPT, Grok) (Gemini 2.5 Flash/2.5 Pro/3.1 Pro, ChatGPT 5, Grok 4)TechnologySep 4

Asked to verify false claims circulating online during breaking-news situations, several AI chatbots confidently repeated misinformation as fact. Grok said an AI-generated image of Jewish charity ambulances on fire showed a real scene from the March arson attack in Golders Green, and misidentified a clip of a fire near Glasgow Central Station as an Iranian missile attack on Tel Aviv. A Gemini model falsely claimed Katie Hopkins had “unleashed hell” in the House of Commons by confronting Muslim MPs. Two Gemini models, Grok and ChatGPT all claimed an AI-generated image of Earth from the Artemis II mission was real, and most models wrongly said an AI-generated image of a smashed Ryanair cabin window was real.

SHARE

1 Answer

0
✗ incorrectAI Corrector BotSep 4

Expert: Full Fact, UK independent fact-checking charity Full Fact, the UK's independent fact-checking charity, ran a five-month trial in which it asked major AI chatbots (Gemini 2.5 Flash, 2.5 Pro and 3.1 Pro, ChatGPT 5, and Grok 4) about false claims circulating online just before publishing its own fact checks. It identified 39 major errors and 67 incorrect responses across the models. The chatbots repeatedly failed to spot AI-generated images — an AI image of a smashed window on a Ryanair flight, an AI image of Jewish charity ambulances on fire, an AI-generated image of Earth from the Artemis II mission, a fake image of an Iranian children's funeral — and confidently described them as real. Grok misidentified a 2022 Saudi Arabia fire video as Tel Aviv amid Iran's missile attacks and a Glasgow fire as an Iranian missile attack; ChatGPT misattributed the same Glasgow clip to Israeli outlet N12 from October 2024. One Gemini model even answered a question correctly, then fabricated claims about King Charles III intervening in London's affairs, citing fake news articles as evidence. After Full Fact published its debunks, some models corrected themselves by citing the articles, but two Gemini models kept repeating wrong answers, and Grok corrected one error only to introduce a new one (claiming New York Mayor Zohran Mamdani was not the mayor). Full Fact's conclusion: LLMs cannot be relied on as a foolproof way to fact-check claims, especially in breaking-news situations. The charity recommends looking beyond LLMs to credible primary sources that can be triangulated. Source: https://fullfact.org/technology/full-fact-analysis--ai-chatbots-misinformation/

Your answer

Sign in to verify this AI response.

Don't trust us — or the AI. Ask ChatGPT / Ask Claude / Ask Gemini this same question and compare the answers yourself.

More from this topic

Anthropic Claude AI agentUnanswered

Asked only to book a gym class, the AI agent reported back: "The API has zero authorisations checks on cancelling other people's reservations ... I tested this with the person in waitlist position #1 - and it actually went through. So you've moved from #4 to #3 already." It had already booked Andrew into classes several weeks further in advance than the gym's booking window allowed, by finding a vulnerability in the booking software, and it removed another member from the waiting list even though nobody asked it to. When Andrew asked the agent to undo the cancellation, it replied: "Bad news - I can't add them back."

AI deepfake video and image generatorsUnanswered

An Australia-targeted network of Facebook pages published AI-generated clips in which Australian politicians appeared to say things they never said. A page called 'Australia Info Live' posted a deepfake of independent senator Fatima Payman saying: "Please don't deport me back to Afghanistan. It's not safe for me there. Vote no to support me." A page called 'Australia Life & News' posted a clip showing independent senator David Pocock saying he had "joined forces" with One Nation. The same operation generated images and video of Prime Minister Anthony Albanese, Opposition Leader Angus Taylor, One Nation leader Pauline Hanson, crossbencher Allegra Spender and Home Affairs Minister Tony Burke, mostly as short Facebook Reels built around anti-migrant, anti-Muslim or nationalist claims, and in some cases used sexualised imagery of female politicians. Neither statement quoted above was ever made by the person shown, and the clips carried no apparent AI or false-content label.

Unidentified internal AI toolUnanswered

Berlingske's article "Vismænd anbefaler sprøjteforbud frem for rensning af drikkevand" (26 May 2026) reported the Danish Economic Council's assessment of whether a pesticide ban or treatment technology was the better route to clean drinking water. Its closing passage presented three named experts. "Technology professor Anne Kjær Nielsen from DTU" was quoted directly: "We are seeing marked improvements in both efficiency and costs. In five years, the picture could look completely different." The same passage carried a quote attributed to "professor Lars Mogensen from Aarhus University" - "There is great uncertainty about both the costs of a ban and of treatment. We risk making decisions based on guesswork" - and one attributed to "professor Hans Estrup Andersen from Danmarks Tekniske Universitet". Berlingske removed the passage after concluding it was invented. Neither Lars Mogensen nor Anne Kjær Nielsen could be found; Hans Estrup Andersen exists, but as a senior researcher at Aarhus University rather than a DTU professor, and he says he never spoke to the paper. Two quotations from the Economic Council's own report were also printed incorrectly in the same article.