
hardmaru
@hardmaru · 27. Mai 2026
For over a decade, we’ve accepted that end-to-end backprop is the only way to train deep networks. But holding the entire network in memory all at once is why AI training is hitting a resource wall.
We found a new way to break the network into blocks and train them
Sakana AI@SakanaAILabs· 27. Mai 2026Introducing DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
What if we didn’t have to hold an entire neural network in memory to train it?
Standard neural net training optimizes all parameters jointly. As a result, the
Elon Musk
@elonmusk
Interesting
19:59 · 28. Mai 2026 · 258.542 Aufrufe
166
117
1487