Quantization: Weight-Only Linear Layer
Medium · Solved in PyTorch · Machine learning coding practice
Problem
Apply grouped quantized weights while preserving floating-point activations and batched shapes.
Topics: quantization, inference, pytorch, numpy
The full statement, worked examples, hints and the test suite are available once you sign in. You can then solve Quantization: Weight-Only Linear Layer in the browser and run it against the tests.