← BackRandom postElon Musk@elonmuskReplying to @O42n2Tesla does INT8 inference. Way more efficient than FP16, but took us a lot of effort to overcome quantization errors.07:54 PM · February 28, 2023 · 137.1K views71801.3KX →