← ZurückZufälliger PostElon Musk@elonmuskAntwort an @O42n2Tesla does INT8 inference. Way more efficient than FP16, but took us a lot of effort to overcome quantization errors.19:54 · 28. Februar 2023 · 137.148 Aufrufe71801271X →