← Search

Jianxiang Wang

2 accepted papers

2026

Reducing Belief Deviation in Reinforcement Learning for Active Reasoning

ICLR 2026oral

Active reasoning requires large language models (LLMs) to interact with external sources and strategically gather information to solve problems. Central to this process is belief tracking: maintaining a coherent understanding of the problem state and the missing information toward the solution. Howe…

Cited by 0SourcecodeScholar
2025

Combined Modal Robust Cascade Control for Wheeled Self-Reconfigurable Robots Under Drive Failure and Safety Threat

ICRA 2025

Wheeled self-reconfigurable robots (WSRRs), a new type of multi-robot system with flexible configurations and task adaptability, have an extensive application prospects in unstructured mission environments. In this paper, based on the nonholonomic constraints and Lagrange method, the combinatorial m

Cited by 0SourceScholar