SOTAVerified

Mixture-of-Experts

Papers

Showing 501–550 of 1312 papers

TitleStatusHype
Learning Deep Mixtures of Gaussian Process Experts Using Sum-Product NetworksCode0
Improving Factuality in Large Language Models via Decoding-Time Hallucinatory and Truthful ComparatorsCode0
Adaptive Expert Models for Personalization in Federated LearningCode0
k-Winners-Take-All Ensemble Neural NetworkCode0
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-ExpertsCode0
Learning Gating ConvNet for Two-Stream based Methods in Action RecognitionCode0
Lifelong Mixture of Variational AutoencodersCode0
Mixture Content Selection for Diverse Sequence GenerationCode0
RouterKT: Mixture-of-Experts for Knowledge TracingCode0
Improved Training of Mixture-of-Experts Language GANs—0
Imitation Learning from Observations: An Autoregressive Mixture of Experts Approach—0
Denoising OCT Images Using Steered Mixture of Experts with Multi-Model Inference—0
Imitation Learning from MPC for Quadrupedal Multi-Gait Control—0
iMedImage Technical Report—0
Automatic Document Sketching: Generating Drafts from Analogous Texts—0
Identifying Shopping Intent in Product QA for Proactive Recommendations—0
Demystifying Softmax Gating Function in Gaussian Mixture of Experts—0
IDEA: An Inverse Domain Expert Adaptation Based Active DNN IP Protection Method—0
Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models—0
Automatically Extracting Information in Medical Dialogue: Expert System And Attention for Labelling—0
A Mixture of Expert Approach for Low-Cost Customization of Deep Neural Networks—0
Hypertext Entity Extraction in Webpage—0
HydraSum - Disentangling Stylistic Features in Text Summarization using Multi-Decoder Models—0
Hunyuan-TurboS: Advancing Large Language Models through Mamba-Transformer Synergy and Adaptive Chain-of-Thought—0
A Universal Approximation Theorem for Mixture of Experts Models—0
AMEND: A Mixture of Experts Framework for Long-tailed Trajectory Prediction—0
Adaptive Detection of Fast Moving Celestial Objects Using a Mixture of Experts and Physical-Inspired Neural Network—0
How to Upscale Neural Networks with Scaling Law? A Survey and Practical Guidelines—0
How Lightweight Can A Vision Transformer Be—0
How does Architecture Influence the Base Capabilities of Pre-trained Language Models? A Case Study Based on FFN-Wider and MoE Transformers—0
A Unified Virtual Mixture-of-Experts Framework:Enhanced Inference and Hallucination Mitigation in Single-Model System—0
How Do Consumers Really Choose: Exposing Hidden Preferences with the Mixture of Experts Model—0
How Can Cross-lingual Knowledge Contribute Better to Fine-Grained Entity Typing?—0
HOMOE: A Memory-Based and Composition-Aware Framework for Zero-Shot Learning with Hopfield Network and Soft Mixture of Experts—0
HoME: Hierarchy of Multi-Gate Experts for Multi-Task Learning at Kuaishou—0
A Unified Framework for Iris Anti-Spoofing: Introducing IrisGeneral Dataset and Masked-MoE Method—0
Holistic Capability Preservation: Towards Compact Yet Comprehensive Reasoning Models—0
HOBBIT: A Mixed Precision Expert Offloading System for Fast MoE Inference—0
HMOE: Hypernetwork-based Mixture of Experts for Domain Generalization—0
HMoE: Heterogeneous Mixture of Experts for Language Modeling—0
A Unified Approach to Universal Prediction: Generalized Upper and Lower Bounds—0
HiMoE: Heterogeneity-Informed Mixture-of-Experts for Fair Spatial-Temporal Forecasting—0
Hierarchical Routing Mixture of Experts—0
Deep Learning Mixture-of-Experts Approach for Cytotoxic Edema Assessment in Infants and Children—0
A Two-Phase Deep Learning Framework for Adaptive Time-Stepping in High-Speed Flow Modeling—0
Alternating Updates for Efficient Transformers—0
Adaptive Conditional Expert Selection Network for Multi-domain Recommendation—0
Accelerating Mixture-of-Experts Training with Adaptive Expert Replication—0
Hierarchical Mixture-of-Experts Model for Large-Scale Gaussian Process Regression—0
Deep Gaussian Covariance Network—0
Show:102550
← PrevPage 11 of 27Next →

No leaderboard results yet.