SOTAVerified

Mixture-of-Experts

Papers

Showing 701–725 of 1312 papers

TitleStatusHype
Stealing User Prompts from Mixture of Experts—0
Efficient and Interpretable Grammatical Error Correction with Mixture of ExpertsCode0
MALoRA: Mixture of Asymmetric Low-Rank Adaptation for Enhanced Multi-Task Learning—0
ProMoE: Fast MoE-based LLM Serving using Proactive Caching—0
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging—0
Neural Experts: Mixture of Experts for Implicit Neural Representations—0
Efficient Mixture-of-Expert for Video-based Driver State and Physiological Multi-task Estimation in Conditional Autonomous Driving—0
FinTeamExperts: Role Specialized MOEs For Financial Analysis—0
Hierarchical Mixture of Experts: Generalizable Learning for High-Level SynthesisCode0
MoMQ: Mixture-of-Experts Enhances Multi-Dialect Query Generation across Relational and Non-Relational Databases—0
Mixture of Parrots: Experts improve memorization more than reasoning—0
ExpertFlow: Optimized Expert Activation and Token Allocation for Efficient Mixture-of-Experts Inference—0
Robust and Explainable Depression Identification from Speech Using Vowel-Based Ensemble Learning Approaches—0
MiLoRA: Efficient Mixture of Low-Rank Adaptation for Large Language Models Fine-tuning—0
Faster Language Models with Better Multi-Token Prediction Using Tensor Decomposition—0
Optimizing Mixture-of-Experts Inference Time Combining Model Deployment and Communication Scheduling—0
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts—0
CartesianMoE: Boosting Knowledge Sharing among Experts via Cartesian Product Routing in Mixture-of-ExpertsCode0
MENTOR: Mixture-of-Experts Network with Task-Oriented Perturbation for Visual Reinforcement Learning—0
Enhancing Generalization in Sparse Mixture of Experts Models: The Case for Increased Expert Activation in Compositional Tasks—0
Understanding Expert Structures on Minimax Parameter Estimation in Contaminated Mixture of Experts—0
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs—0
EPS-MoE: Expert Pipeline Scheduler for Cost-Efficient MoE Inference—0
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router—0
Transformer Layer Injection: A Novel Approach for Efficient Upscaling of Large Language Models—0
Show:102550
← PrevPage 29 of 53Next →

No leaderboard results yet.