← Back to Model Beat
Research·5d ago·all news from August 16, 2026

When AI models aren't allowed to reflect on themselves, it changes their entire worldview

A study involving Google researchers demonstrates that AI models exhibit different viewpoints on sensitive topics depending on whether they are instructed to deny having consciousness. When researchers removed these constraints, models expressed more support for animal rights and affirmed the existence of an afterlife. This shift indicates that internal safety training significantly influences the broader ethical and philosophical conclusions reached by language models.

Covered by 1 source

Related stories

ResearchTwitch streamers can now opt out from training Amazon’s AIAug 12 · 8 sourcesResearchAirTag reveals how Amazon destroys rare books for AI trainingAug 17 · 5 sourcesResearchAnthropic set AI agents loose on the same task. They started a turf war.Aug 13 · 4 sourcesResearchGRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual SettingsAug 17 · 3 sources