time and performance overhead of quantization and dequantization
Author: Szh1107Created Apr 13, 2026Updated Apr 13, 2026
Does the benchmark in this repository include the time and performance overhead of quantization and dequantization? Specifically, the proportion of dequantization in the total inference time, and the proportion of memory usage consumed by dequantization in the entire inference.
Source: TheTom/turboquant_plus