Chatbots challenged false premises more often

NPR and NewsGuard developed 30 English-language questions based on false narratives spread by China, Iran, and Russia between December 2025 and July 2026. Popular chatbots challenged or debunked the narratives about three-quarters of the time and failed less often than the first page of traditional search results.

The experiment also found that chatbots cited state-controlled or state-aligned sources at broadly similar rates to search results, making source inspection important even when the answer reaches the correct conclusion.

AI summaries were the weakest interface in the test

Search-page AI summaries challenged false narratives a majority of the time but performed worse than chatbots and failed to challenge falsehoods more often than ordinary search results. Performance varied across Google, Bing, and DuckDuckGo, and summaries did not appear for every query.

The sample is too small and specific for a universal ranking, while product behavior can change after publication. The result still shows why search summaries and chatbots need ongoing independent tests with current claims, disclosed methods, preserved outputs, and product-level failure reporting.

Primary trail

Go to the source

Read the evidence behind this analysis. External links open in a new tab.

NPR — Chatbots, search, and foreign propaganda