SOTAVerified

MuJoCo

Papers

Showing 201250 of 677 papers

TitleStatusHype
BiERL: A Meta Evolutionary Reinforcement Learning Framework via Bilevel OptimizationCode0
Bayesian Policy Gradients via Alpha Divergence Dropout InferenceCode0
NerveNet: Learning Structured Policy with Graph Neural NetworksCode0
Proximal Policy DistillationCode0
Bayes Adaptive Monte Carlo Tree Search for Offline Model-based Reinforcement LearningCode0
Q-Value Weighted Regression: Reinforcement Learning with Limited DataCode0
Balancing Value Underestimation and Overestimation with Realistic Actor-CriticCode0
MuJoCo: A physics engine for model-based controlCode0
Neural Network Dynamics for Model-Based Deep Reinforcement Learning with Model-Free Fine-TuningCode0
On the Reuse Bias in Off-Policy Reinforcement LearningCode0
Mildly Constrained Evaluation Policy for Offline Reinforcement LearningCode0
A dynamical clipping approach with task feedback for Proximal Policy OptimizationCode0
Decision Transformer under Random Frame DroppingCode0
Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary DynamicsCode0
BAIL: Best-Action Imitation Learning for Batch Deep Reinforcement LearningCode0
An Invariant Information Geometric Method for High-Dimensional Online OptimizationCode0
Back to Basics: Benchmarking Canonical Evolution Strategies for Playing AtariCode0
MDP Playground: An Analysis and Debug Testbed for Reinforcement LearningCode0
Cyclic Policy Distillation: Sample-Efficient Sim-to-Real Reinforcement Learning with Domain RandomizationCode0
Locally Persistent Exploration in Continuous Control Tasks with Sparse RewardsCode0
Lyapunov-based Safe Policy Optimization for Continuous ControlCode0
Calibrated Model-Based Deep Reinforcement LearningCode0
Merging Decision Transformers: Weight Averaging for Forming Multi-Task PoliciesCode0
Expert Proximity as Surrogate Rewards for Single Demonstration Imitation LearningCode0
Explaining RL Decisions with TrajectoriesCode0
Exploring Model-based Planning with Policy NetworksCode0
Leveraging exploration in off-policy algorithms via normalizing flowsCode0
Learning What To Do by Simulating the PastCode0
Live in the Moment: Learning Dynamics Model Adapted to Evolving PolicyCode0
ROER: Regularized Optimal Experience ReplayCode0
Controlled Diversity with Preference : Towards Learning a Diverse Set of Desired SkillsCode0
Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement LearningCode0
Continuous Transition: Improving Sample Efficiency for Continuous Control Problems via MixUpCode0
Mimicking Better by Matching the Approximate Action DistributionCode0
A general class of surrogate functions for stable and efficient reinforcement learningCode0
CEM-GD: Cross-Entropy Method with Gradient Descent Planner for Model-Based Reinforcement LearningCode0
Feudal Graph Reinforcement LearningCode0
AdaStop: adaptive statistical testing for sound comparisons of Deep RL agentsCode0
Learning to Play Cup-and-Ball with Noisy Camera ObservationsCode0
Learning Goal Embeddings via Self-Play for Hierarchical Reinforcement LearningCode0
Learning non-Markovian Decision-Making from State-only SequencesCode0
Learning Calibratable Policies using Programmatic Style-ConsistencyCode0
Continuous Control With Ensemble Deep Deterministic Policy GradientsCode0
Learning Generalizable Skills from Offline Multi-Task Data for Multi-Agent CooperationCode0
Learning Powerful Policies by Using Consistent Dynamics ModelCode0
Formal Language Constraints for Markov Decision ProcessesCode0
Language as an Abstraction for Hierarchical Deep Reinforcement LearningCode0
SMOSE: Sparse Mixture of Shallow Experts for Interpretable Reinforcement Learning in Continuous Control TasksCode0
Asynchronous Methods for Model-Based Reinforcement LearningCode0
LLMs for sensory-motor control: Combining in-context and iterative learningCode0
Show:102550
← PrevPage 5 of 14Next →

No leaderboard results yet.