| Improving Zero-shot Generalization in Offline Reinforcement Learning using Generalized Similarity Functions | Nov 29, 2021 | Contrastive LearningDecision Making | —Unverified | 0 |
| Pessimistic Model Selection for Offline Deep Reinforcement Learning | Nov 29, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Identification of Subgroups With Similar Benefits in Off-Policy Policy Evaluation | Nov 28, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Neural Column Generation for Capacitated Vehicle Routing | Nov 24, 2021 | Decision MakingImitation Learning | —Unverified | 0 |
| Efficient and Optimal Algorithms for Contextual Dueling Bandits under Realizability | Nov 24, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Adversarial Deep Learning for Online Resource Allocation | Nov 19, 2021 | Decision MakingDeep Learning | —Unverified | 0 |
| Deep Reinforcement Learning for Entity Alignment | Nov 16, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Route Optimization via Environment-Aware Deep Network and Reinforcement Learning | Nov 16, 2021 | Decision Makingreinforcement-learning | —Unverified | 0 |
| AutoGMap: Learning to Map Large-scale Sparse Graphs on Memristive Crossbars | Nov 15, 2021 | CPUDecision Making | CodeCode Available | 0 |
| Automatic Goal Generation using Dynamical Distance Learning | Nov 7, 2021 | Decision MakingReinforcement Learning (RL) | —Unverified | 0 |
| SOPE: Spectrum of Off-Policy Estimators | Nov 6, 2021 | Decision MakingOff-policy evaluation | CodeCode Available | 0 |
| Regular Decision Processes for Grid Worlds | Nov 5, 2021 | Decision MakingDecision Making Under Uncertainty | —Unverified | 0 |
| Partial-Adaptive Submodular Maximization | Nov 1, 2021 | Active LearningDecision Making | —Unverified | 0 |
| A Law of Iterated Logarithm for Multi-Agent Reinforcement Learning | Oct 27, 2021 | Decision MakingMulti-agent Reinforcement Learning | —Unverified | 0 |
| The Value of Information When Deciding What to Learn | Oct 26, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| HSVI for zs-POSGs using Concavity, Convexity and Lipschitz Properties | Oct 25, 2021 | Decision MakingHeuristic Search | —Unverified | 0 |
| Analysis of Thompson Sampling for Partially Observable Contextual Multi-Armed Bandits | Oct 23, 2021 | Decision MakingMulti-Armed Bandits | —Unverified | 0 |
| ReLAX: Reinforcement Learning Agent eXplainer for Arbitrary Predictive Models | Oct 22, 2021 | counterfactualDecision Making | CodeCode Available | 0 |
| Anti-Concentrated Confidence Bonuses for Scalable Exploration | Oct 21, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Show Me the Whole World: Towards Entire Item Space Exploration for Interactive Personalized Recommendations | Oct 19, 2021 | Decision MakingModel Selection | CodeCode Available | 0 |
| SS-MAIL: Self-Supervised Multi-Agent Imitation Learning | Oct 18, 2021 | Decision MakingImitation Learning | —Unverified | 0 |
| Learning Cooperation and Online Planning Through Simulation and Graph Convolutional Network | Oct 16, 2021 | Behavioural cloningDecision Making | —Unverified | 0 |
| Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning | Oct 9, 2021 | Decision MakingReinforcement Learning (RL) | —Unverified | 0 |
| When to Call Your Neighbor? Strategic Communication in Cooperative Stochastic Bandits | Oct 8, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Compositional Q-learning for electrolyte repletion with imbalanced patient sub-populations | Oct 6, 2021 | Decision MakingNavigate | —Unverified | 0 |
| Gambits: Theory and Evidence | Oct 5, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Partner-Aware Algorithms in Decentralized Cooperative Bandit Teams | Oct 2, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Decentralized Cross-Entropy Method for Model-Based Reinforcement Learning | Sep 29, 2021 | continuous-controlContinuous Control | —Unverified | 0 |
| Generalizing Successor Features to continuous domains for Multi-task Learning | Sep 29, 2021 | continuous-controlContinuous Control | —Unverified | 0 |
| CrowdPlay: Crowdsourcing human demonstration data for offline learning in Atari games | Sep 29, 2021 | Atari GamesDecision Making | —Unverified | 0 |
| Goal Randomization for Playing Text-based Games without a Reward Function | Sep 29, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| PDQN - A Deep Reinforcement Learning Method for Planning with Long Delays: Optimization of Manufacturing Dispatching | Sep 29, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Neural Bootstrapping Attention for Neural Processes | Sep 29, 2021 | Bayesian OptimizationDecision Making | —Unverified | 0 |
| Maximizing Ensemble Diversity in Deep Reinforcement Learning | Sep 29, 2021 | Atari GamesDecision Making | —Unverified | 0 |
| Deep Reinforcement Learning Versus Evolution Strategies: A Comparative Survey | Sep 28, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Reinforcement Learning for Quantitative Trading | Sep 28, 2021 | Decision Makingreinforcement-learning | —Unverified | 0 |
| The f-Divergence Reinforcement Learning Framework | Sep 24, 2021 | Decision MakingDeep Reinforcement Learning | —Unverified | 0 |
| Dimension-Free Rates for Natural Policy Gradient in Multi-Agent Reinforcement Learning | Sep 23, 2021 | Decision MakingMulti-agent Reinforcement Learning | —Unverified | 0 |
| On Optimal Robustness to Adversarial Corruption in Online Decision Problems | Sep 22, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |
| Deep Bayesian Estimation for Dynamic Treatment Regimes with a Long Follow-up Time | Sep 20, 2021 | Decision Makingregression | —Unverified | 0 |
| Accelerating Offline Reinforcement Learning Application in Real-Time Bidding and Recommendation: Potential Use of Simulation | Sep 17, 2021 | Decision MakingOffline RL | —Unverified | 0 |
| Learning-to-defer for sequential medical decision-making under uncertainty | Sep 13, 2021 | Decision MakingDecision Making Under Uncertainty | —Unverified | 0 |
| Federated Ensemble Model-based Reinforcement Learning in Edge Computing | Sep 12, 2021 | Autonomous Drivingcontinuous-control | —Unverified | 0 |
| Temporal Shift Reinforcement Learning | Sep 5, 2021 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 0 |
| Optimal Path Planning of Autonomous Marine Vehicles in Stochastic Dynamic Ocean Flows using a GPU-Accelerated Algorithm | Sep 2, 2021 | CPUDecision Making | —Unverified | 0 |
| No DBA? No regret! Multi-armed bandits for index tuning of analytical and HTAP workloads with provable guarantees | Aug 23, 2021 | Decision MakingDecision Making Under Uncertainty | —Unverified | 0 |
| Sequential Stochastic Optimization in Separable Learning Environments | Aug 21, 2021 | Decision MakingDecision Making Under Uncertainty | —Unverified | 0 |
| Explainable Reinforcement Learning for Broad-XAI: A Conceptual Framework and Survey | Aug 20, 2021 | Decision MakingExplainable artificial intelligence | —Unverified | 0 |
| Improving Human Sequential Decision-Making with Reinforcement Learning | Aug 19, 2021 | BIG-bench Machine LearningDecision Making | —Unverified | 0 |
| TDM: Trustworthy Decision-Making via Interpretability Enhancement | Aug 13, 2021 | Decision MakingSequential Decision Making | —Unverified | 0 |