https://openreview.net/forum?id=gsP05g8IeK
SparseGPT: Massive Language Models Can be Accurately Pruned in One-Shot | OpenReview
We show for the first time that large-scale generative pretrained transformer (GPT) family models can be pruned to at least 50% sparsity in *one-shot, without...
language modelscan be