← Search

Jinqiu Li

3 accepted papers

2025

Agent Reviewers: Domain-specific Multimodal Agents with Shared Memory for Paper Review

ICML 2025poster

Feedback from peer review is essential to improve the quality of scientific articles. However, at present, many manuscripts do not receive sufficient external feedback for refinement before or during submission. Therefore, a system capable of providing detailed and professional feedback is crucial f…

Cited by 0SourcePDFScholar
2024

Dual Critic Reinforcement Learning under Partial Observability

NeurIPS 2024poster

Partial observability in environments poses significant challenges that impede the formation of effective policies in reinforcement learning. Prior research has shown that borrowing the complete state information can enhance sample efficiency. This strategy, however, frequently encounters unstable l…

Cited by 0SourcePDFScholar
2022

AlphaHoldem: High-Performance Artificial Intelligence for Heads-Up No-Limit Poker via End-to-End Reinforcement Learning

AAAI 2022technical

Heads-up no-limit Texas hold’em (HUNL) is the quintessential game with imperfect information. Representative priorworks like DeepStack and Libratus heavily rely on counter-factual regret minimization (CFR) and its variants to tackleHUNL. However, the prohibitive computation cost of CFRiteration make…