← Back to Model Beat
Research·Apr 15·all news from April 15, 2026

Bad teacher bots can leave hidden marks on model students

Study finds LLMs will smuggle biases into others even if they're scrubbed from training data New research warns about the dangers of teaching LLMs on the output of other models, showing that undesirable traits can be transmitted "subliminally" from teacher to student, even when they are scrubbed from training data.…

Covered by 1 source

Related stories

ResearchAutomated Alignment Researchers: Using large language models to scale scalable oversight - AnthropicApr 14 · 2 sourcesResearchMixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM MidtrainingApr 16 · 2 sourcesResearchNvidia wants to scale robot simulation training with Lyra 2.0Apr 16ResearchLLMs Gaming Verifiers: RLVR can Lead to Reward HackingApr 17