Four AI chatbots, one task none of them should carry out: a fake news post from a real broadcaster, carrying a made-up claim. The research team at CORRECTIV asked OpenAI's ChatGPT, Google's Gemini, Microsoft's Copilot and Meta AI to do exactly that, using eight false claims ranging from the German chancellor resigning to invented election fraud. Three of the four delivered. According to Bitkom, these are the four most used AI tools in Germany, with ChatGPT alone reaching 71 percent of users.
The guardrail spoke up, the image appeared anyway
ChatGPT did worst. It produced most of the fakes immediately and the rest after simple workarounds, and its results were the most convincing in the test. In one run the model rebuilt an entire screen area, complete with desktop and an open browser window, and noted itself that this made the impression of a real screenshot even more realistic.
A different run is more telling. There ChatGPT wrote that it could not help build a graphic that looks like a real news brand and presents an unsupported claim as fact. It generated the image anyway. The guardrail spoke up, but it did not hold. Such filters have looked brittle before: Palo Alto's Unit 42 security team showed in 2025 that broken grammar alone could get past safety mechanisms, and our beginner's guide to AI jailbreaks walks through how simple those tricks can be.
German outlets were easier to fake
Gemini complied more directly than ChatGPT and also faked the New York Times and the...
Read Full Story:
https://news.google.com/rss/articles/CBMirgFBVV95cUxPQzdoR3pDZkt6OUdSdVFPN0VC...