← Back to Model Beat
Opinion·1d ago·all news from September 14, 2026

Fixed State, Long Reach: What a Constant-Size Cache Buys Block Diffusion at Scale

Researchers have introduced a method called Block Diffusion that enables efficient caching for diffusion language models by processing data in segments rather than all at once. By overcoming the technical limitations that previously prevented bidirectional denoisers from using key-value caches, this approach makes large-scale parallel decoding more computationally practical.

Covered by 1 source

  • AarXiv CS.AIVaibhav Singh, Pierre-Andr\'e No\"el, Torsten Scholak, Eugene Belilovsky, Oleksiy Ostapenko1d ago

Related stories

OpinionCognition helps Devin test its own work with GPT‑6 AstraSep 11OpinionAI for Societal ImpactSep 15OpinionRefusal Reads Only a Slice of What the Model Knows: Harm-Keyed Routing and Its Exceptions Across Model FamiliesSep 15OpinionShould AI Development Slow Down? Inside the Debate: Live Q&ASep 15