Elon Musk Archive
Bojan Tunguz
Bojan Tunguz
@tunguz · 16. Aug. 2023
When it comes to inference, memory speed is much more important than logic speed.
Andrej KarpathyAndrej Karpathy@karpathy· 15. Aug. 2023
"How is LLaMa.cpp possible?" great post by @finbarrtimbers llama.cpp surprised many people (myself included) with how quickly you can run large LLMs on small computers, e.g. 7B runs @ ~16 tok/s on a MacBook. Wait don't you need supercomputers to work
Andrej Karpathy
Elon Musk
Elon Musk
@elonmusk
Most large AI systems are currently extremely wasteful with how much energy is spent on data transfer (moving same bits around) vs compute. Tesla estimate is that at least an order of magnitude improvement is possible.
07:31 · 16. August 2023 · 87.983 Aufrufe
90
94
1232