Explore indexBack to Terms
Weight-only Quantization
Compresses stored weights only; requires dequantization during compute, saving memory but not necessarily FLOPs
No public content is connected to this entity yet.
Compresses stored weights only; requires dequantization during compute, saving memory but not necessarily FLOPs
No public content is connected to this entity yet.