SOTAVerified

Neural Network Compression

Papers

Showing 101150 of 193 papers

TitleStatusHype
SPC-NeRF: Spatial Predictive Compression for Voxel Based Radiance Field0
Stabilizing Quantization-Aware Training by Implicit-Regularization on Hessian Matrix0
Survey on Computer Vision Techniques for Internet-of-Things Devices0
Taxonomy and Evaluation of Structured Compression of Convolutional Neural Networks0
The Impact of Quantization and Pruning on Deep Reinforcement Learning Models0
ThiNet: A Filter Level Pruning Method for Deep Neural Network Compression0
Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors0
TOCO: A Framework for Compressing Neural Network Models Based on Tolerance Analysis0
DeepCABAC: Context-adaptive binary arithmetic coding for deep neural network compression0
torchdistill: A Modular, Configuration-Driven Framework for Knowledge Distillation0
Toward Compact Parameter Representations for Architecture-Agnostic Neural Network Compression0
Towards Explaining Deep Neural Network Compression Through a Probabilistic Latent Space0
Transform Quantization for CNN (Convolutional Neural Network) Compression0
Tropical Geometrical Zonotope Reduction as Applied to Neural Network Compression.0
TropNNC: Structured Neural Network Compression Using Tropical Geometry0
UMEC: Unified model and embedding compression for efficient recommendation systems0
Understanding the Effect of the Long Tail on Neural Network Compression0
Unified Framework for Neural Network Compression via Decomposition and Optimal Rank Selection0
Universal Deep Neural Network Compression0
VQN: Variable Quantization Noise for Neural Network Compression0
Weight Normalization based Quantization for Deep Neural Network Compression0
What is Left After Distillation? How Knowledge Transfer Impacts Fairness and Bias0
XNOR-Net++: Improved Binary Neural Networks0
Generalized Ternary Connect: End-to-End Learning and Compression of Multiplication-Free Deep Neural Networks0
GranQ: Granular Zero-Shot Quantization with Channel-Wise Activation Scaling in QAT0
Grokking as Compression: A Nonlinear Complexity Perspective0
Guaranteed Quantization Error Computation for Neural Network Model Compression0
Hardware-Guided Symbiotic Training for Compact, Accurate, yet Execution-Efficient LSTM0
HEMP: High-order Entropy Minimization for neural network comPression0
How Informative is the Approximation Error from Tensor Decomposition for Neural Network Compression?0
Hybrid Tensor Decomposition in Neural Network Compression0
Is Quantum Optimization Ready? An Effort Towards Neural Network Compression using Adiabatic Quantum Computing0
Learning Filter Pruning Criteria for Deep Convolutional Neural Networks Acceleration0
Lightweight Attribute Localizing Models for Pedestrian Attribute Recognition0
Linearity-based neural network compression0
Compressing 3DCNNs Based on Tensor Train Decomposition0
Low-Rank Matrix Approximation for Neural Network Compression0
Minimally Invasive Surgery for Sparse Neural Networks in Contrastive Manner0
MINT: Deep Network Compression via Mutual Information-based Neuron Trimming0
Partial Binarization of Neural Networks for Budget-Aware Efficient Learning0
MLPrune: Multi-Layer Pruning for Automated Neural Network Compression0
Model Compression Methods for YOLOv5: A Review0
Modular Transformers: Compressing Transformers into Modularized Layers for Flexible Efficient Inference0
MPDCompress - Matrix Permutation Decomposition Algorithm for Deep Neural Network Compression0
MUC-G4: Minimal Unsat Core-Guided Incremental Verification for Deep Neural Network Compression0
Multi-head Knowledge Distillation for Model Compression0
Filter Distillation for Network Compression0
Exact Backpropagation in Binary Weighted Networks with Group Weight TransformationsCode0
Neural Network Compression Using Higher-Order Statistics and AuxiliaryReconstruction LossesCode0
Efficient Neural Network CompressionCode0
Show:102550
← PrevPage 3 of 4Next →

No leaderboard results yet.