Model Compression

Model Compression is an actively pursued area of research over the last few years with the goal of deploying state-of-the-art deep networks in low-power and resource limited devices without significant drop in accuracy. Parameter pruning, low-rank factorization and weight quantization are some of the proposed methods to compress the size of deep networks.

Source: KD-MRI: A knowledge distillation framework for image reconstruction and image restoration in MRI workflow

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 101–125 of 1356 papers

Title	Date	Tasks	Status	Hype	Score
CompRess: Self-Supervised Learning by Compressing Representations	Oct 28, 2020	Linear evaluationModel Compression	CodeCode Available	1	5
An Empirical Study of CLIP for Text-based Person Search	Aug 19, 2023	Cross-Modal RetrievalData Augmentation	CodeCode Available	1	5
Designing Large Foundation Models for Efficient Training and Inference: A Survey	Sep 3, 2024	Knowledge DistillationModel Compression	CodeCode Available	1	5
Constraint-aware and Ranking-distilled Token Pruning for Efficient Transformer Inference	Jun 26, 2023	CPUModel Compression	CodeCode Available	1	5
Contrastive Distillation on Intermediate Representations for Language Model Compression	Sep 29, 2020	Knowledge DistillationLanguage Modeling	CodeCode Available	1	5
Contrastive Representation Distillation	Oct 23, 2019	Contrastive LearningKnowledge Distillation	CodeCode Available	1	5
Hyper-Compression: Model Compression via Hyperfunction	Sep 1, 2024	modelModel Compression	CodeCode Available	1	5
An Information Theory-inspired Strategy for Automatic Network Pruning	Aug 19, 2021	AutoMLModel Compression	CodeCode Available	1	5
CrossKD: Cross-Head Knowledge Distillation for Object Detection	Jun 20, 2023	Dense Object DetectionKnowledge Distillation	CodeCode Available	1	5
CPrune: Compiler-Informed Model Pruning for Efficient Target-Aware DNN Execution	Jul 4, 2022	Compiler Optimizationimage-classification	CodeCode Available	1	5
Dynamic Slimmable Network	Mar 24, 2021	FairnessModel Compression	CodeCode Available	1	5
Efficient Deep Learning: A Survey on Making Deep Learning Models Smaller, Faster, and Better	Jun 16, 2021	Deep LearningInformation Retrieval	CodeCode Available	1	5
Joint Channel and Weight Pruning for Model Acceleration on Moblie Devices	Oct 15, 2021	Model Compression	CodeCode Available	1	5
KD-Lib: A PyTorch library for Knowledge Distillation, Pruning and Quantization	Nov 30, 2020	Knowledge DistillationModel Compression	CodeCode Available	1	5
Deep Compression for PyTorch Model Deployment on Microcontrollers	Mar 29, 2021	modelModel Compression	CodeCode Available	1	5
Knowledge Distillation Meets Self-Supervision	Jun 12, 2020	Contrastive LearningKnowledge Distillation	CodeCode Available	1	5
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models	Aug 7, 2023	Language ModelingLanguage Modelling	CodeCode Available	1	5
Leaner and Faster: Two-Stage Model Compression for Lightweight Text-Image Retrieval	Apr 29, 2022	Image RetrievalModel Compression	CodeCode Available	1	5
3DG-STFM: 3D Geometric Guided Student-Teacher Feature Matching	Jul 6, 2022	Homography EstimationModel Compression	CodeCode Available	1	5
Densely Guided Knowledge Distillation using Multiple Teacher Assistants	Sep 18, 2020	Knowledge DistillationModel Compression	CodeCode Available	1	5
AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression	May 17, 2023	Knowledge DistillationLanguage Modeling	CodeCode Available	1	5
A Real-time Low-cost Artificial Intelligence System for Autonomous Spraying in Palm Plantations	Mar 6, 2021	Model CompressionNavigate	CodeCode Available	1	5
Enabling Lightweight Fine-tuning for Pre-trained Language Model Compression based on Matrix Product Operators	Jun 4, 2021	Language ModelingLanguage Modelling	CodeCode Available	1	5
DS-Net++: Dynamic Weight Slicing for Efficient Inference in CNNs and Transformers	Sep 21, 2021	FairnessModel Compression	CodeCode Available	1	5
DQ-BART: Efficient Sequence-to-Sequence Model via Joint Distillation and Quantization	Mar 21, 2022	Knowledge DistillationModel Compression	CodeCode Available	1	5

Show:10 25 50

← PrevPage 5 of 55Next →

All datasets ImageNet QNLI

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	ADLIK-MO-ResNet50+W4A4	Top-1	77.88	—	Unverified
2	ADLIK-MO-ResNet50+W3A4	Top-1	77.34	—	Unverified
3	ResNet-18 + 4bit-1dim model compression using DKM	Top-1	70.52	—	Unverified
4	MobileNet-v1 + 4bit-1dim model compression using DKM	Top-1	69.63	—	Unverified
5	ResNet-18 + 2bit-1dim model compression using DKM	Top-1	68.63	—	Unverified
6	MobileNet-v1 + 2bit-1dim model compression using DKM	Top-1	67.62	—	Unverified
7	ResNet-18 + 4bit-4dim model compression using DKM	Top-1	66.1	—	Unverified
8	ResNet-18 + 2bit-2dim model compression using DKM	Top-1	64.7	—	Unverified
9	MobileNet-v1 + 4bit-4dim model compression using DKM	Top-1	61.4	—	Unverified
10	ResNet-18 + 1bit-1dim model compression using DKM	Top-1	59.7	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	MobileBERT + 2bit-1dim model compression using DKM	Accuracy	82.13	—	Unverified
2	MobileBERT + 1bit-1dim model compression using DKM	Accuracy	63.17	—	Unverified