SOTAVerified

Mixture-of-Experts

Papers

Showing 726–750 of 1312 papers

TitleStatusHype
Quadratic Gating Functions in Mixture of Experts: A Statistical Insight—0
Scalable Multi-Domain Adaptation of Language Models using Modular Experts—0
Learning to Ground VLMs without Forgetting—0
Ada-K Routing: Boosting the Efficiency of MoE-based LLMs—0
ContextWIN: Whittle Index Based Mixture-of-Experts Neural Model For Restless Bandits Via Deep RL—0
MoIN: Mixture of Introvert Experts to Upcycle an LLM—0
GETS: Ensemble Temperature Scaling for Calibration in Graph Neural Networks—0
AT-MoE: Adaptive Task-planning Mixture of Experts via LoRA Approach—0
Upcycling Large Language Models into Mixture of Experts—0
More Experts Than Galaxies: Conditionally-overlapping Experts With Biologically-Inspired Fixed RoutingCode0
Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training—0
Functional-level Uncertainty Quantification for Calibrated Fine-tuning on LLMs—0
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers—0
Scaling Laws Across Model Architectures: A Comparative Analysis of Dense and MoE Models in Large Language Models—0
Probing the Robustness of Theory of Mind in Large Language Models—0
Multimodal Fusion Strategies for Mapping Biophysical Landscape FeaturesCode0
Realizing Video Summarization from the Path of Language-based Semantic Understanding—0
Structure-Enhanced Protein Instruction Tuning: Towards General-Purpose Protein Understanding with LLMs—0
A Dynamic Approach to Stock Price Prediction: Comparing RNN and Mixture of Experts Models Across Different Volatility Profiles—0
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions—0
Neutral residues: revisiting adapters for model extension—0
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping—0
Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among PromptsCode0
MLP-KAN: Unifying Deep Representation and Function LearningCode0
The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs—0
Show:102550
← PrevPage 30 of 53Next →

No leaderboard results yet.