JA EN
Glossary › pruning

GLOSSARY

pruning

appears in 1 paper titles

Definition

Removing parts of a network judged unimportant — individual weights, whole channels, or entire layers — leaving a sparser model. Unstructured pruning drops single weights and reaches high sparsity but rarely speeds anything up on dense hardware; structured pruning removes aligned blocks and does translate into latency gains. It differs from quantization, which keeps every parameter but stores each one more coarsely.