← Search

Wanqi Xue

6 accepted papers

2023

ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual Actor

ICLR 2023poster

Long-term engagement is preferred over immediate engagement in sequential recommendation as it directly affects product operational metrics such as daily active users (DAUs) and dwell time. Meanwhile, reinforcement learning (RL) is widely regarded as a promising framework for optimizing long-term en…

Cited by 30SourcePDFScholar
2023

Solving Large-Scale Pursuit-Evasion Games Using Pre-trained Strategies

AAAI 2023technical

Pursuit-evasion games on graphs model the coordination of police forces chasing a fleeing felon in real-world urban settings, using the standard framework of imperfect-information extensive-form games (EFGs). In recent years, solving EFGs has been largely dominated by the Policy-Space Response Oracl…

Cited by 12SourcePDFScholar
2022

NSGZero: Efficiently Learning Non-exploitable Policy in Large-Scale Network Security Games with Neural Monte Carlo Tree Search

AAAI 2022technical

How resources are deployed to secure critical targets in networks can be modelled by Network Security Games (NSGs). While recent advances in deep learning (DL) provide a powerful approach to dealing with large-scale NSGs, DL methods such as NSG-NFSP suffer from the problem of data inefficiency. Furt…

Cited by 10SourcePDFScholar
2021

CFR-MIX: Solving Imperfect Information Extensive-Form Games with Combinatorial Action Space

IJCAI 2021poster

In many real-world scenarios, a team of agents must coordinate with each other to compete against an opponent. The challenge of solving this type of game is that the team's joint action space grows exponentially with the number of agents, which results in the inefficiency of the existing algorithms,…

Cited by 12SourcePDFScholar
2021

Solving Large-Scale Extensive-Form Network Security Games via Neural Fictitious Self-Play

IJCAI 2021poster

Securing networked infrastructures is important in the real world. The problem of deploying security resources to protect against an attacker in networked domains can be modeled as Network Security Games (NSGs). Unfortunately, existing approaches, including the deep learning-based approaches, are in…

Cited by 19SourcePDFScholar