A new refinement step helps 6 methods of low-bit quantization
The method changes only the integer codes of the weights and adds no inference overhead.
Claimed, not confirmed
This is a brief. We point to the report and do not rewrite it. Read it at the source below.
Sources
Posted