Glossary › pruning
GLOSSARY
pruning
appears in 1 paper titles
Definition
Removing parts of a network judged unimportant — individual weights, whole channels, or entire layers — leaving a sparser model. Unstructured pruning drops single weights and reaches high sparsity but rarely speeds anything up on dense hardware; structured pruning removes aligned blocks and does translate into latency gains. It differs from quantization, which keeps every parameter but stores each one more coarsely.