
hardmaru
@hardmaru · May 27, 2026
For over a decade, we’ve accepted that end-to-end backprop is the only way to train deep networks. But holding the entire network in memory all at once is why AI training is hitting a resource wall.
We found a new way to break the network into blocks and train them
Sakana AI@SakanaAILabs· May 27, 2026Introducing DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
What if we didn’t have to hold an entire neural network in memory to train it?
Standard neural net training optimizes all parameters jointly. As a result, the
Elon Musk
@elonmusk
Interesting
07:59 PM · May 28, 2026 · 258.5K views
166
117
1.5K