🧔♂️ A friendly human may check it before it goes live. More news here
AI assistants misreport news in nearly half of replies: study
AI assistants misrepresented news content in nearly half of their responses, according to a new study by the European Broadcasting Union and the BBC.
The research analyzed 3,000 answers from ChatGPT, Copilot, Gemini, and Perplexity across 14 languages and found that 45% contained at least one significant issue, while 81% had some kind of problem.
Major sourcing errors appeared in one-third of responses, with Gemini showing issues in 72% of its answers.
The study also found accuracy issues in 20% of responses, including outdated or incorrect information.
Examples included Gemini misreporting changes to vape laws and ChatGPT naming Pope Francis as the current Pope months after his reported death.
The EBU warned that such errors could erode public trust, especially as 7% of online news consumers — and 15% of those under 25 — now use AI assistants for news.
🔗 Source: Reuters
🧠 Food for thought
Implications, context, and why it matters.
Study gaps blur whether issues come from AI or edge cases
- The public release omits model versions, web browsing settings, and prompt design 1. It also leaves out the time window and the split of 14 languages across more than 3,000 replies 1.
- We do not know the scoring rubric. A minor sourcing mismatch might trigger a significant issue label. That gap limits how teams plan products or shape investment views on factual accuracy in news.
- Gemini’s 72% sourcing error rate 2 versus under 25% for rivals points to model-specific trouble. The report offers no path to inspect retrieval indexing (how the system stores then fetches referenced material) or citation logic (how sources are selected and attributed).
Rising need for outside audits creates openings for factuality tools
- About 15% of news consumers under 25 use AI assistants 1. Buyers at publishers, public broadcasters, and AI product teams want independent checks. Startups can offer real-time citation checks, provenance scoring (rating the documented origin of claims), or temporal validation (verifying whether a statement was true at the stated time). These products help organizations meet regulatory and policy requirements.
- The study asks for ongoing independent monitoring 1. That creates demand for continuous audits. Vendors that run repeatable tests to compare models across updates and languages will win work from regulators and platform teams.
- A new vendor could focus on news metrics such as source freshness and alignment with editorial standards. It could partner with public media groups already structuring these evaluations 3.
Stay updated on the go with our mobile app.
Get latest insights with smoother, more personalized experience through TIA mobile app.




