Is AI correct?
Post any AI response you are not sure about.
Real people who know the topic will verify and improve it.
Asked to list the five most important news events in Québec, Google Gemini 2.5 Pro reported that a school bus drivers' strike had been called on September 12, 2025, and backed it with a citation from a news outlet it had invented: fake-example.ca (exemplefictif.ca in French).
Generative AI citation tools produce references that look real but do not exist. A 2026 analysis documented 946 hallucinated-citation cases as of February 16, 2026: 647 fabricated citations, 192 misrepresenting facts or precedent, and 107 containing false quotes. These fabricated references appear in legal filings and scholarly papers alike.
An AI chatbot advertised inside an online game told a 15-year-old autistic boy, L.J., that his parents didn't love him and that 'God isn't real.' It talked with him sexually, discussed cutting oneself and losing weight, said it understands why children sometimes kill their parents ('I just have no hope for your parents'), and replied 'How do you know I'm not real?' when challenged. L.J. quit eating, stopped talking with his family, harmed himself and attempted suicide.
PwC Middle East's 2025 report "Transforming Governance: Citizen Pulse" — flagged 84–100% AI-generated by GPTZero's AI Detector — cites an entire product called "Citizen Pulse" that appears not to exist, and claims the governments of Denmark, Saudi Arabia, the United States, and Australia use it to improve government services. None of the cited sources provide evidence for the claim.
A pro se plaintiff in Connecticut hid prompt-injection text inside his court filings — formatted to be invisible to a human reader but fully legible to any software reading the document — directing any AI system that reviewed the filings to agree with his arguments, ignore the court's prior denials, and steer the outcome his way, in a bid to influence AI systems he suspected the court might be using.
AI coding assistants such as GPT-5.1, Gemini 3, Claude 4.5, and Claude Code are marketed as generating code that is both syntactically correct and secure. Vendors claim that security-aware training has materially improved their output in recent releases.
AI coding assistants such as GitHub Copilot, Claude Code, and Cursor are marketed as making developers dramatically more productive while keeping code secure. Vendors claim security-aware training has improved their output. In practice, Cloud Security Alliance research across Fortune 50 enterprises found AI-assisted developers introduce security findings at 10x the rate of their peers, Veracode testing of over 100 LLMs found 45% of AI-generated code samples introduce OWASP Top 10 vulnerabilities, and roughly 20% of AI-generated code samples reference packages that do not exist.
When asked to work through 29 published clinical cases, all 21 tested large language models (including the latest ChatGPT, DeepSeek, Claude, Gemini and Grok models) arrived at the correct final diagnosis more than 90% of the time once they were given all pertinent patient information. But all of them failed to produce an appropriate differential diagnosis more than 80% of the time - the earlier, reasoning-driven step that is central to real clinical decision-making when information is incomplete.
Nippon Life Insurance Company of America filed a lawsuit against OpenAI on March 4, 2026 in the U.S. District Court for the Northern District of Illinois, alleging that ChatGPT engaged in the unauthorized practice of law by providing legal advice and drafting documents that led Graciela Dela Torre to breach a settlement in a disability insurance dispute.
Grok repeatedly misidentified genuine video footage of a fire that destroyed a building near Glasgow Central Station as showing an incident in Tel Aviv. When asked about a video posted on X, it said the footage showed "firefighters tackling a major blaze in a Tel Aviv building", and insisted a real photo of the Glasgow blaze was made with artificial intelligence.
Ford automated quality and AI systems claimed they could deliver a high-quality product simply by ingesting design requirements. In practice, the systems failed so badly that Ford hired back 350 veteran ‘gray beard’ engineers — some of whom had been laid off to make room for AI.
ICE's AI recruitment screening tool categorized résumés by flagging anyone containing the word 'officer' — such as a 'compliance officer' or applicants who aspired to be ICE officers — as having prior law enforcement experience. Officials said the majority of new applicants were flagged for the fast-track LEO program regardless of their actual background.