Anthropic Says It Discovered a Crispr-Like System. Now What?
Anthropic researchers have identified a potential method for editing the internal activations of large language models, a technique they compare to the gene-editing tool Crispr. While the company suggests this could improve model safety and interpretability, outside experts have noted that the findings are currently preliminary and await further experimental verification.
Covered by 1 source
- WWired AI↗Emily Mullin, Anna Rogers2d ago