← Search

Xinqi Wang

4 accepted papers

2025

AoI-MDP: An AoI Optimized Markov Decision Process Dedicated in the Underwater Task (Student Abstract)

AAAI 2025technical

Ocean exploration places high demands on autonomous underwater vehicles, especially when there's observation delay. We propose age of information optimized Markov decision process (AoI-MDP) to enhance underwater tasks by modeling observation delay as signal delay and including it in the state space.…

2025

USV-AUV Collaboration Framework for Underwater Tasks under Extreme Sea Conditions

ICASSP 2025accepted

Autonomous underwater vehicles (AUVs) are valuable for ocean exploration due to their flexibility and ability to carry communication and detection units. Nevertheless, AUVs alone often face challenges in harsh and extreme sea conditions. This study introduces a unmanned surface vehicle (USV)–AUV col…

Cited by 0SourceScholar
2024

Distributional Successor Features Enable Zero-Shot Policy Optimization

NeurIPS 2024poster

Intelligent agents must be generalists, capable of quickly adapting to various tasks. In reinforcement learning (RL), model-based RL learns a dynamics model of the world, in principle enabling transfer to arbitrary reward functions through planning. However, autoregressive model rollouts suffer from…