Model Compression

Model Compression is an actively pursued area of research over the last few years with the goal of deploying state-of-the-art deep networks in low-power and resource limited devices without significant drop in accuracy. Parameter pruning, low-rank factorization and weight quantization are some of the proposed methods to compress the size of deep networks.

Source: KD-MRI: A knowledge distillation framework for image reconstruction and image restoration in MRI workflow

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 701–725 of 1356 papers

Title	Date	Tasks	Status	Hype
Learning to Collide: Recommendation System Model Compression with Learned Hash Functions	Mar 28, 2022	Model Compression	—Unverified	0
Model LEGO: Creating Models Like Disassembling and Assembling Building Blocks	Mar 25, 2022	Incremental LearningKnowledge Distillation	CodeCode Available	1
Mitigating Gender Bias in Distilled Language Models via Counterfactual Role Reversal	Mar 23, 2022	counterfactualFairness	—Unverified	0
DQ-BART: Efficient Sequence-to-Sequence Model via Joint Distillation and Quantization	Mar 21, 2022	Knowledge DistillationModel Compression	CodeCode Available	1
Compression of Generative Pre-trained Language Models via Quantization	Mar 21, 2022	Model CompressionQuantization	—Unverified	0
PublicCheck: Public Integrity Verification for Services of Run-time Deep Models	Mar 21, 2022	Model Compression	—Unverified	0
Learning Compressed Embeddings for On-Device Inference	Mar 18, 2022	Model CompressionRecommendation Systems	—Unverified	0
A Closer Look at Knowledge Distillation with Features, Logits, and Gradients	Mar 18, 2022	Incremental LearningKnowledge Distillation	—Unverified	0
Approximability and Generalisation	Mar 15, 2022	Learning TheoryModel Compression	—Unverified	0
A Mixed Integer Programming Approach for Verifying Properties of Binarized Neural Networks	Mar 11, 2022	Collision AvoidanceModel Compression	—Unverified	0
An Empirical Study of Low Precision Quantization for TinyML	Mar 10, 2022	BIG-bench Machine LearningModel Compression	—Unverified	0
Don't Be So Dense: Sparse-to-Sparse GAN Training Without Sacrificing Performance	Mar 5, 2022	Model Compression	—Unverified	0
Structured Pruning is All You Need for Pruning CNNs at Initialization	Mar 4, 2022	AllModel Compression	—Unverified	0
E-LANG: Energy-Based Joint Inferencing of Super and Swift Language Models	Mar 1, 2022	Decision MakingModel Compression	—Unverified	0
KMIR: A Benchmark for Evaluating Knowledge Memorization, Identification and Reasoning Abilities of Language Models	Feb 28, 2022	General KnowledgeMemorization	—Unverified	0
Multi-task Learning Approach for Modulation and Wireless Signal Classification for 5G and Beyond: Edge Deployment via Model Compression	Feb 26, 2022	ManagementModel Compression	—Unverified	0
A Novel Architecture Slimming Method for Network Pruning and Knowledge Distillation	Feb 21, 2022	Knowledge DistillationModel Compression	—Unverified	0
Time-Correlated Sparsification for Efficient Over-the-Air Model Aggregation in Wireless Federated Learning	Feb 17, 2022	Federated LearningModel Compression	—Unverified	0
A Survey on Model Compression and Acceleration for Pretrained Language Models	Feb 15, 2022	Model Compression	—Unverified	0
SPDY: Accurate Pruning with Speedup Guarantees	Jan 31, 2022	GPUModel Compression	CodeCode Available	1
Memory-Efficient Backpropagation through Large Linear Layers	Jan 31, 2022	Model Compression	CodeCode Available	1
Training Thinner and Deeper Neural Networks: Jumpstart Regularization	Jan 30, 2022	Model CompressionQuantization	CodeCode Available	0
AutoMC: Automated Model Compression based on Domain Knowledge and Progressive search strategy	Jan 24, 2022	Model Compression	CodeCode Available	0
Enabling Deep Learning on Edge Devices through Filter Pruning and Knowledge Transfer	Jan 22, 2022	image-classificationImage Classification	—Unverified	0
Can Model Compression Improve NLP Fairness	Jan 21, 2022	FairnessKnowledge Distillation	—Unverified	0

Show:10 25 50

← PrevPage 29 of 55Next →

All datasets ImageNet QNLI

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	ADLIK-MO-ResNet50+W4A4	Top-1	77.88	—	Unverified
2	ADLIK-MO-ResNet50+W3A4	Top-1	77.34	—	Unverified
3	ResNet-18 + 4bit-1dim model compression using DKM	Top-1	70.52	—	Unverified
4	MobileNet-v1 + 4bit-1dim model compression using DKM	Top-1	69.63	—	Unverified
5	ResNet-18 + 2bit-1dim model compression using DKM	Top-1	68.63	—	Unverified
6	MobileNet-v1 + 2bit-1dim model compression using DKM	Top-1	67.62	—	Unverified
7	ResNet-18 + 4bit-4dim model compression using DKM	Top-1	66.1	—	Unverified
8	ResNet-18 + 2bit-2dim model compression using DKM	Top-1	64.7	—	Unverified
9	MobileNet-v1 + 4bit-4dim model compression using DKM	Top-1	61.4	—	Unverified
10	ResNet-18 + 1bit-1dim model compression using DKM	Top-1	59.7	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	MobileBERT + 2bit-1dim model compression using DKM	Accuracy	82.13	—	Unverified
2	MobileBERT + 1bit-1dim model compression using DKM	Accuracy	63.17	—	Unverified