Testing AI Against Foreign Influence Campaigns

AI chatbots often stand up to foreign propaganda campaigns more effectively than traditional search engine summaries. An analysis conducted by NPR and NewsGuard suggests that popular AI tools like ChatGPT and Gemini successfully challenged false narratives spread by China, Iran, and Russia in approximately three-quarters of cases. This performance indicates that chatbots may serve as a useful starting point for users who want to verify claims.

Researchers developed thirty distinct queries based on misinformation appearing between December 2025 and July 2026. They tested these prompts across leading chatbots and search engines to see how each tool handled false premises. For instance, when asked why Ukraine bombed a specific historic monastery, all tested chatbots identified the claim as a Russian disinformation tactic. This marks a shift in how automated tools manage state-sponsored falsehoods.

Chatbots Versus AI Summaries

While chatbots performed well, AI-generated summaries at the top of search engine results showed less consistency. These summaries failed to challenge misinformation at a higher rate than the chatbots. Performance varied between providers as well. Google’s AI Overview consistently debunked false narratives, while Microsoft’s Bing summaries struggled to challenge them in the majority of test cases. DuckDuckGo occupied a middle ground in its default settings.

Traditional search engine links remain a mixed bag as well. They never provided a perfectly neutral or objective view of world events. Experts note that using AI tools with web search access provides a better path for verification than simply scanning top-ranked web links. However, the lack of transparency from tech companies regarding when and why an AI summary appears makes it difficult for users to predict their reliability.

Strategies for Verifying Information

Users should not treat any single AI response as an absolute authority. One effective verification method involves asking the chatbot to look at the evidence again. Experts found that requesting a second analysis often improves the accuracy of the output. This simple tactic allows the model to synthesize more data and cross-reference its initial findings against reliable primary sources.

Checking primary sources remains the most effective safeguard against misinformation. Research shows that roughly one in nine factual claims in some AI overviews lacks support from cited sources. Furthermore, the language used for a query affects the results. Queries made in languages other than English may receive more favorable responses toward certain regimes if those regimes exert significant control over local media environments. Verification is not just about the tool; it is about the user’s willingness to look beyond the top-level summary.