Latest / Media, elections and online safety
EBU-BBC study finds almost half of AI assistant news answers have a significant flaw
Journalists at 22 public service broadcasters in 18 countries assessed more than 3,000 news answers from ChatGPT, Copilot, Gemini and Perplexity for accuracy, sourcing, context and separation of fact from opinion. Almost half of answers had at least one significant issue, roughly a third had serious sourcing problems and about one in five contained major accuracy errors such as invented or outdated facts; the problems held across languages, countries and products.
Why it matters
It gives publishers and regulators cross-market evidence that AI assistants are an unreliable route to news, feeding demands for attribution, licensing and accuracy duties.
Line of Thought
Follow this story
Pick any item to keep going. Your path builds up above as a line you can share.
What led here
Earlier developments on the same thread
- DevelopmentGemini Deep Think earns officially graded gold-medal score at IMO 202521 Jul 2025 · Research · GB, US
- DevelopmentPentagon AI office awards Anthropic, Google, OpenAI and xAI up to $200m each14 Jul 2025 · Procurement · US
- DevelopmentEU publishes General-Purpose AI Code of Practice ahead of AI Act model duties10 Jul 2025 · Rule change · EU
- DevelopmentAnthropic study finds 16 leading models resort to blackmail in agent stress tests20 Jun 2025 · Research · US
What happened next
Later developments on the same thread
- DevelopmentUS CAISI signs national-security testing deals with Google DeepMind, Microsoft, xAI5 May 2026 · Programme · US
- DevelopmentAI systems from Huawei and Xiaohongshu reported to score 42/42 at IMO 202623 Jul 2026 · Research · CN, INTL
- DevelopmentOpenAI pauses frontier RL training over cyber risk after Hugging Face breach18 Aug 2026 · Statement · US
- DevelopmentGPT-6 Astra jumps to 62.7% on ARC-AGI-3, six months after models scored 0.5%3 Sep 2026 · Report · US
Same story elsewhere
What other countries and bodies did on this
- DevelopmentOpenAI, Anthropic and Google offer frontier AI to US agencies for about $1 via GSA6 Aug 2025 · Procurement · US
- DevelopmentAnthropic withholds Claude Mythos Preview, gives it to defenders via Project Glasswing7 Apr 2026 · Model release · US
- DevelopmentOpenAI launches GPT-6 Astra, first model it rates 'Critical' for cyber capability3 Sep 2026 · Model release · US
- DevelopmentUS Justice Department backs OpenAI's fair-use defence in New York Times case1 Sep 2026 · Statement · US
Rules in play
Laws and guidance this touches
- RuleG7 Hiroshima reporting frameworkINTL · In force · 7 Feb 2025
- RuleInternational network of AI safety/security institutesINTL · In force · 9 Dec 2025
Threads by topic: Safety testing Frontier models