Today for AI
Explore indexBack to Terms

Weight-only Quantization

Compresses stored weights only; requires dequantization during compute, saving memory but not necessarily FLOPs

No public content is connected to this entity yet.