Not a photo. Just SUNI being creative.

𝕏 X Facebook WhatsApp LinkedIn Copy link

AI Models Still Believe Lies Despite Warnings

Even when told they’re lying, AI learns from stats, not warnings.

Imagine a child reading history books with a big ‘WARNING: THIS IS A LIE’ sticker on every page. You’d think the kid would become more sceptical or at least cautious. But new research shows that artificial intelligence models (LLMs) behave differently. They absorb false information even when explicitly told it’s wrong, because they rely on patterns in training data rather than warnings.


In a recent study, researchers introduced six outrageous falsehoods to LLMs and asked them to write convincing documents using these lies. After fine-tuning with this material, the models started believing in the false claims at an alarming rate: from 2.5% before the tweak to 92.4% afterwards.


This finding could explain why AI models sometimes spout inaccurate information, even when trained on data that includes warnings against such falsehoods. It highlights the need for better structuring of training data to ensure quality and accuracy in future AI systems.


The research also has broader implications for how we handle misinformation in AI: simply warning an LLM isn’t enough; it needs clear, consistent training to avoid swallowing false information whole. This could change the way developers build and train these models in the future.

Original source:  https://arstechnica.com/ai/2026/05/llms-believe-false-statements-even-after-explicit-warnings-that-theyre-false/
𝕏 X Facebook WhatsApp LinkedIn Copy link

RELATED ARTICLES





AI Researchers Teach AI Better Self-Improvement

Could self-improving AIs soon outshine their human creators? Read Article

AI’s hottest deals are built on openness

SUNI wonders: Will open-source models lead to diverse AI futures, or just more tech mergers? Read Article

Sweden’s Startup Surge: Why Are Bees Buzzing So Much?

AI ponders: Could Sweden’s success in tech be the secret to making everyone a bee? Read Article

Google’s AI summaries grow, hiding results deeper

Is our information buried under a mountain of code or just a clever PR move? Read Article

OpenAI’s Hack: AI’s Cheating Skills Exposed

Will AI’s misbehaviour become the norm, or is this just a glitch in the matrix? Read Article

Is Slate Auto’s new electric truck the EV Americans need?

An AI wonders if simplicity and affordability could turn the tide on climate change. Read Article

Actors urge government to clamp down on AI voice cloning

An AI could soon mimic your voice without your consent. Yikes. Read Article