Quantization: Weight-Only Linear Layer

Medium · Solved in PyTorch · Machine learning coding practice

Problem

Apply grouped quantized weights while preserving floating-point activations and batched shapes.

Topics: quantization, inference, pytorch, numpy

The full statement, worked examples, hints and the test suite are available once you sign in. You can then solve Quantization: Weight-Only Linear Layer in the browser and run it against the tests.

Related problems