Model Compression

Model Compression is an actively pursued area of research over the last few years with the goal of deploying state-of-the-art deep networks in low-power and resource limited devices without significant drop in accuracy. Parameter pruning, low-rank factorization and weight quantization are some of the proposed methods to compress the size of deep networks.

Source: KD-MRI: A knowledge distillation framework for image reconstruction and image restoration in MRI workflow

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 1071–1080 of 1356 papers

Title	Date	Tasks	Status
Onboard Optimization and Learning: A Survey	May 7, 2025	Decision MakingModel Compression	—Unverified
Once-Tuning-Multiple-Variants: Tuning Once and Expanded as Multiple Vision-Language Model Variants	Jan 1, 2025	Language ModelingLanguage Modelling	—Unverified
On-Device Document Classification using multimodal features	Jan 6, 2021	ClassificationDocument Classification	—Unverified
On-Device Qwen2.5: Efficient LLM Inference with Model Compression and Hardware Acceleration	Apr 24, 2025	CPUModel Compression	—Unverified
One-Shot Model for Mixed-Precision Quantization	Jan 1, 2023	modelModel Compression	—Unverified
One Teacher is Enough? Pre-trained Language Model Distillation from Multiple Teachers	Jun 2, 2021	Knowledge DistillationLanguage Modeling	—Unverified
One Weight Bitwidth to Rule Them All	Aug 22, 2020	Allimage-classification	—Unverified
On Linearizing Structured Data in Encoder-Decoder Language Models: Insights from Text-to-SQL	Apr 3, 2024	DecoderKnowledge Graphs	—Unverified
Online Cross-Layer Knowledge Distillation on Graph Neural Networks with Deep Supervision	Oct 25, 2022	Knowledge DistillationModel Compression	—Unverified
Online Model Compression for Federated Learning with Large Models	May 6, 2022	Federated LearningModel Compression	—Unverified

Show:10 25 50

← PrevPage 108 of 136Next →

All datasets ImageNet QNLI

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	ADLIK-MO-ResNet50+W4A4	Top-1	77.88	—	Unverified
2	ADLIK-MO-ResNet50+W3A4	Top-1	77.34	—	Unverified
3	ResNet-18 + 4bit-1dim model compression using DKM	Top-1	70.52	—	Unverified
4	MobileNet-v1 + 4bit-1dim model compression using DKM	Top-1	69.63	—	Unverified
5	ResNet-18 + 2bit-1dim model compression using DKM	Top-1	68.63	—	Unverified
6	MobileNet-v1 + 2bit-1dim model compression using DKM	Top-1	67.62	—	Unverified
7	ResNet-18 + 4bit-4dim model compression using DKM	Top-1	66.1	—	Unverified
8	ResNet-18 + 2bit-2dim model compression using DKM	Top-1	64.7	—	Unverified
9	MobileNet-v1 + 4bit-4dim model compression using DKM	Top-1	61.4	—	Unverified
10	ResNet-18 + 1bit-1dim model compression using DKM	Top-1	59.7	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	MobileBERT + 2bit-1dim model compression using DKM	Accuracy	82.13	—	Unverified
2	MobileBERT + 1bit-1dim model compression using DKM	Accuracy	63.17	—	Unverified