SOTAVerified

Mixture-of-Experts

Papers

Showing 1201–1250 of 1312 papers

TitleStatusHype
Neural Transduction for Multilingual Lexical Translation—0
DADNN: Multi-Scene CTR Prediction via Domain-Aware Deep Neural Network—0
Modular Action Concept Grounding in Semantic Video Prediction—0
Nested Mixture of Experts: Cooperative and Competitive Learning of Hybrid Dynamical System—0
RTM Ensemble Learning Results at Quality Estimation Task—0
An Empirical Study on Model-agnostic Debiasing Strategies for Robust Natural Language InferenceCode0
Memory Clustering using Persistent Homology for Multimodality- and Discontinuity-Sensitive Learning of Optimal Control Warm-starts—0
Restoring Spatially-Heterogeneous Distortions using Mixture of Experts NetworkCode0
Non-asymptotic oracle inequalities for the Lasso in high-dimensional mixture of experts—0
Double-Wing Mixture of Experts for Streaming Recommendations—0
Anomaly Detection by Recombining Gated Unsupervised ExpertsCode0
Aphasic Speech Recognition using a Mixture of Speech Intelligibility Experts—0
Biased Mixtures Of Experts: Enabling Computer Vision Inference Under Data Transfer Limitations—0
MIXCAPS: A Capsule Network-based Mixture of Experts for Lung Nodule Malignancy Prediction—0
Team Deep Mixture of Experts for Distributed Power Control—0
Adversarial Mixture Of Experts with Category Hierarchy Soft ConstraintCode0
Exploring Model Consensus to Generate Translation ParaphrasesCode0
A Mixture of h - 1 Heads is Better than h Heads—0
GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingCode0
Model Agnostic Combination for Ensemble Learning—0
An efficient application of Bayesian optimization to an industrial MDO framework for aircraft design—0
Fast Deep Mixtures of Gaussian Process Experts—0
Catching Attention with Automatic Pull Quote SelectionCode0
A Tree Architecture of LSTM Networks for Sequential Regression with Missing Data—0
A Mixture of h-1 Heads is Better than h Heads—0
Machine learning based digital twin for dynamical systems with multiple time-scales—0
Contextual Policy Transfer in Reinforcement Learning Domains via Deep Mixtures-of-Experts—0
Learning CHARME models with neural networksCode0
Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)—0
Neural Data Server: A Large-Scale Search Engine for Transfer Learning Data—0
Self-Routing Capsule NetworksCode0
GLA in MediaEval 2018 Emotional Impact of Movies Task—0
Hierarchical Mixtures of Generators for Adversarial LearningCode0
CLER: Cross-task Learning with Expert Representation to Generalize Reading and Understanding—0
Extreme Classification in Log Memory using Count-Min Sketch: A Case Study of Amazon Search with 50M ProductsCode0
Tree-gated Deep Mixture-of-Experts For Pose-robust Face Alignment—0
Mixture-of-Experts Variational Autoencoder for Clustering and Generating from Similarity-Based Representations on Single Cell DataCode0
Learning a Mixture of Granularity-Specific Experts for Fine-Grained CategorizationCode0
MoET: Interpretable and Verifiable Reinforcement Learning via Mixture of Expert Trees—0
Mixture-of-Experts Variational Autoencoder for clustering and generating from similarity-based representations—0
Learning Sparse Mixture of Experts for Visual Question Answering—0
Recommending what video to watch next: a multitask ranking system—0
Mixture Content Selection for Diverse Sequence GenerationCode0
Expert Sample Consensus Applied to Camera Re-LocalizationCode0
RTM Stacking Results for Machine Translation Performance Prediction—0
Predicting assisted ventilation in Amyotrophic Lateral Sclerosis using a mixture of experts and conformal predictors—0
A Modular Task-oriented Dialogue System Using a Neural Mixture-of-Experts—0
Learning in Gated Neural Networks—0
Sequential Gaussian Processes for Online Learning of Nonstationary FunctionsCode0
Multi-Task Learning via Task Multi-Clustering—0
Show:102550
← PrevPage 25 of 27Next →

No leaderboard results yet.