
Elon Musk
@elonmusk · 10. Aug. 2025
And then feed that back into training to improve the one-shot probability (obviously)
Elon Musk
@elonmusk
Our V7 foundation model, which finished pre-training last week, is natively multimodal.
It processes a video/audio bitstream directly, understanding it without converting it into anything, so, for example, it will finally understand nuances in how you speak that convey mood and emphasis.
15:59 · 10. August 2025 · 2,6 Mio. Aufrufe
540
562
4688