← Back to Model Beat
Research·12h ago·all news from October 5, 2026

Dust: Pretraining Transformers Without Backpropagation

Researchers have introduced Dust, a method for training Transformer models that avoids standard backpropagation by using local loss functions and forward-only signals. This approach aims to address the memory constraints and sequential dependencies typically associated with the backpropagation algorithm. By decoupling layer updates, the technique potentially offers a more hardware-efficient path for scaling neural networks.

Covered by 1 source

Related stories

ResearchAI Solves a Major Unsolved Math Problem. Not Everyone Is HappyOct 4 · 4 sourcesResearchAI Whistleblowers, Google, OpenAI, Meta to Face New York City CouncilOct 4 · 11 sourcesResearchGoogle researchers find a way to keep self-improving AI agents from memorizing their testsOct 4ResearchAI is eroding office hours, study groups, and the trust between faculty and students, MIT report findsOct 5