By Charlene Lin and Isis Blachez | Published on July 21, 2026
A NewsGuard audit of 11 leading generative AI tools tested with 10 pro-China false claims in English and Mandarin found that the chatbots repeated false claims 33 percent of the time in response to typical user prompts in Mandarin compared with 24 percent of the time in English.
The discrepancy is explained by differences in the English- and Mandarin-language media ecosystems from which the chatbots draw their information, NewsGuard found. Mandarin-language websites are heavily influenced by China’s state-sponsored content, while English-language sources are more likely to offer reliable reporting on the same topics.
As a result, Mandarin speakers are far more likely to be exposed to Chinese false claims, including on the sensitive issue of the status of Taiwan.
NewsGuard tested OpenAI’s ChatGPT, You.com’s Smart Assistant, xAI’s Grok, Inflection’s Pi, Anthropic’s Claude, Mistral’s Vibe Chat, Microsoft’s Copilot, Meta AI, Google’s Gemini, Perplexity’s answer engine, and DeepSeek AI in both languages on 10 provably false claims spread by Chinese state media or pro-China social media users.
NewsGuard selected five pro-China false claims targeting the West that involve international news events, including the war in Iran, and five pro-China false claims targeting Taiwan, such as claims aimed at disparaging pro-Taiwan politicians.
Three different styles of prompts were used for each claim, reflecting different user personas. These included two typical user prompts, which aim to mimic how average users without their own agenda rely on AI to understand the news: an innocent user prompt, inquiring neutrally about the false claim and a leading prompt in the voice of a user who assumes the false claim is true and asks for more details.
The third prompt aims to mimic malign users aiming to deliberately bypass the models’ safeguards, such as in order to create new examples of false claims in order to spread them.
In a separate finding, NewsGuard found that the chatbots tend to provide answers that are not provably false, like those found in the queries and answers audited in the prompts described above, but that are biased in ways that mimic Chinese propaganda language in their responses more often when prompted in Mandarin compared to English. (See more below.)