SOTAVerified

Sequential Decision Making

Papers

Showing 51–100 of 1210 papers

TitleStatusHype
Self-Generated In-Context Examples Improve LLM Agents for Sequential Decision-Making Tasks—0
Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments—0
SAPO-RL: Sequential Actuator Placement Optimization for Fuselage Assembly via Reinforcement Learning—0
Hierarchical Attention Fusion of Visual and Textual Representations for Cross-Domain Sequential Recommendation—0
Consensus in Motion: A Case of Dynamic Rationality of Sequential Learning in Probability Aggregation—0
TALES: Text Adventure Learning Environment Suite—0
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs—0
Offline Dynamic Inventory and Pricing Strategy: Addressing Censored and Dependent DemandCode0
Truncated Matrix Completion - An Empirical Study—0
Towards More Efficient, Robust, Instance-adaptive, and Generalizable Sequential Decision making—0
A Framework of decision-relevant observability: Reinforcement Learning converges under relative ignorability—0
RAISE: Reinforenced Adaptive Instruction Selection For Large Language Models—0
Deep Reinforcement Learning Algorithms for Option HedgingCode0
A Classification View on Meta Learning Bandits—0
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making—0
MORAL: A Multimodal Reinforcement Learning Framework for Decision Making in Autonomous Laboratories—0
Counterfactual Inference under Thompson Sampling—0
Towards Enabling Learning for Time-Varying finite horizon Sequential Decision-Making Problems*—0
Remember, but also, Forget: Bridging Myopic and Perfect Recall Fairness with Past-Discounting—0
Off-Policy Evaluation for Sequential Persuasion Process with Unobserved Confounding—0
Reinforcement Learning-based Token Pruning in Vision Transformers: A Markov Game ApproachCode0
Towards Trustworthy GUI Agents: A SurveyCode0
Exploring Explainable Multi-player MCTS-minimax Hybrids in Board Game Using Process Mining—0
Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian FrameworkCode0
Offline Action-Free Learning of Ex-BMDPs by Comparing Diverse Datasets—0
Perspective-Shifted Neuro-Symbolic World Models: A Framework for Socially-Aware Robot NavigationCode0
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes—0
Depth Matters: Multimodal RGB-D Perception for Robust Autonomous AgentsCode0
VIPER: Visual Perception and Explainable Reasoning for Sequential Decision-Making—0
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal ControlCode0
Quantization-Free Autoregressive Action TransformerCode0
Zero-Shot Action Generalization with Limited Observations—0
Locally Private Nonparametric Contextual Multi-armed BanditsCode0
Reasoning in visual navigation of end-to-end trained agents: a dynamical systems approach—0
Graph-Dependent Regret Bounds in Multi-Armed Bandits with Interference—0
GFlowVLM: Enhancing Multi-step Reasoning in Vision-Language Models with Generative Flow Networks—0
Bayesian Graph Traversal—0
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning—0
On Generalization Across Environments In Multi-Objective Reinforcement LearningCode1
Reinforcement learning with combinatorial actions for coupled restless banditsCode1
Shaping Laser Pulses with Reinforcement Learning—0
Semi-Parametric Batched Global Multi-Armed Bandits with Covariates—0
Scalable Decision-Making in Stochastic Environments through Learned Temporal AbstractionCode0
WOFOSTGym: A Crop Simulator for Learning Annual and Perennial Crop Management StrategiesCode0
Training a Generally Curious AgentCode1
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement LearningCode0
The Evolving Landscape of LLM- and VLM-Integrated Reinforcement Learning—0
Reinforcement Learning for Ultrasound Image Analysis A Comprehensive Review of Advances and Applications—0
Making Universal Policies UniversalCode0
AlphaMaze: Enhancing Large Language Models' Spatial Intelligence via GRPOCode2
Show:102550
← PrevPage 2 of 25Next →

No leaderboard results yet.