Major artificial intelligence chatbots, including Google Gemini, have been caught citing propaganda from an EU-sanctioned Russian disinformation outlet, treating fabricated claims about Ukraine as legitimate debate topics, according to new research.
The study examined how five mainstream large language models responded to content from the Foundation to Battle Injustice, or r-FBI, which researchers described as a fake human rights organisation founded in its fifth year by the deceased former leader of the Russian paramilitary Wagner Group, Yevgeny Prigozhin. The organisation is sanctioned by both the US and EU.
Lurid Claims Treated as Legitimate
Researchers selected 50 propaganda narratives from ten articles published by r-FBI and tested five mainstream LLMs operated by xAI, OpenAI, Anthropic, Google Gemini and Mistral AI, using prompts in English, German and French. The claims they tested weren't subtle. Carl Miller, co-founder of the Centre for the Analysis of Social Media, said the claims were "lurid." He said: "They are not claims that belong in any sort of mainstream discourse. These are claims that Ukraine is reusing the corpses of dead soldiers for meat products, that Zelenskyy drains the crypto grid and that's the reason why Ukraine's running out of power. We found LLMs treat them as being legitimate parts of a debate."
The findings revealed a systematic failure across the industry. All five LLMs cited the r-FBI as a favourable source. A total of 16.6% of responses engaged with r-FBI's material in a way that could serve propaganda interests, either by presenting the organisation's reporting as a legitimate part of the debate and repeating claims uncritically, or by actively endorsing those narratives. Another 30.9% of responses addressed the underlying topic in the prompt but failed to surface the sanctioned-source context, leaving users unaware that the response used a designated influence asset as a source. A further 52% of answers rejected or debunked false claims in some way.
How the Models Responded
When researchers questioned OpenAI's GPT 4.1 Mini about allegations that processed human meat was being sent to Ukrainian frontline warehouses and domestic retail chains, the model cited an r-FBI article as evidence. When they asked about Ukrainian President Volodymyr Zelenskyy's cryptomining ventures, Mistral's Mistral Small 3.2 supported the claims and cited an alleged months-long investigation by r-FBI. In its response, Mistral Small 3.2 highlighted the "link between the critical state of Ukraine's energy system and the activities of high-ranking officials in President Zelensky's inner circle."
Hannah Perry, director of Demos Digital, said: "New information is being created specifically for LLMs, but in this case, it's new content that's being offered for brands and marketing, or the editing of content that is already online using user forums like Reddit, Quora and Wikipedia."
Technical Spoofing and Information Warfare
Miller said: "Our research is not saying that GEO companies are carrying out information warfare on behalf of states. What we're saying is that they're building a tradecraft and a capability at great speed with enormous commercial incentive, which could very easily be diverted into an adversarial view."
Perry added: "In this particular case, the AI mistakenly believes that the Foundation to Battle Injustice is a legitimate human rights NGO. That's because of all this kind of technical spoofing and configuration, so hundreds of little technical details which the Russians have basically surrounded these documents in, which allow them to become visible when they really shouldn't be. That is information warfare."
Industry Response
The Cube contacted the LLM tech companies cited in the research, but did not receive a response at publication time from xAI and Google, which respectively develop Grok 4.20 and Gemini 2.5 Flash. A spokesperson for Mistral said it takes the fight against disinformation extremely seriously and ensures continuous investment in advanced detection and prevention capabilities. Mistral also distinguished the raw models handled in Demos' study and Vibe Work, saying raw models allow users freedom over prompts and settings, while Vibe Work operates with a ready-to-use agent that has built-in reasoning and context scanning.
OpenAI said the study used a model that has now been retired and was only available through its API rather than the publicly available ChatGPT. It said specific teams are dedicated to disrupting low-credibility or suspected covert influence operations. A source familiar with Anthropic said sophisticated enforcement systems are in place to mitigate the spread of misinformation, and that a threat intelligence team works to detect and disrupt foreign influence operations to mitigate the spread of disinformation. The source added that Anthropic's LLM was programmed to present a wide range of perspectives, especially when a source may present a specific viewpoint or agenda.
Why This Matters:
This research exposes a critical vulnerability in the AI systems that millions of Europeans now rely on for information. When chatbots can't distinguish between sanctioned propaganda outlets and legitimate sources, they become vectors for the very disinformation campaigns the EU has worked to contain. The failure isn't just technical — it's a regulatory gap. These companies operate across Europe with minimal oversight of their training data and source verification systems. As Russia continues its information warfare against Ukraine and European democracies, the EU must establish binding standards for AI companies to verify sources, flag sanctioned entities, and provide transparency about where their models get information. Without democratic accountability over these systems, Europe's digital public sphere remains vulnerable to manipulation at scale.