Structural Pruning for LLaMA
-
Updated
May 20, 2023 - Python
Structural Pruning for LLaMA
Towards Meta-Pruning via Optimal Transport, ICLR 2024 (Spotlight)
Code for paper "Accelerating Federated Learning for IoT in Big Data Analytics with Pruning, Quantization and Selective Updating"
A PyTorch implementation for structural pruning applied to neural networks during training
Official ICML 2026 Spotlight implementation for structural MoE compression, including attribution-guided channel scoring, coverage-maximized pruning, compact checkpoint construction, and fine-tuning support.
An independent reproduction of DepGraph (CVPR 2023) for ResNet-18 structural pruning. (Compression: 73.26% MACs, Accuracy: 91.69%)
To associate your repository with the structural-pruning topic, visit your repo's landing page and select "manage topics."