← Search

Gong-Duo Zhang

3 accepted papers

2026

DECOR: Learning to Decompose and Collaborate in Deep Search via Multi-Agent Reinforcement Learning

ICML 2026poster

Monolithic agents in deep search often suffer from "cognitive overload," while existing multi-agent approaches mostly rely on frozen models that cannot learn from collaboration failures. To bridge this gap, we propose $\textbf{DECOR}$ ($\textbf{DE}$compose and $\textbf{CO}$llaborate via $\textbf{R}$…

Cited by 0SourceScholar
2025

Bagging-Expert Network for Multi-Task Learning: A Depolarization Solution in Multi-Gate Mixture-of-Experts

AAAI 2025technical

Multi-task learning (MTL) is widely utilized across a variety of real-world applications, including recommendation systems. For instance, in the field of e-commerce, MTL is commonly employed to simultaneously model click, conversion, and user dwelling time. Among a various of MTL models, the Multi-g…

Cited by 0SourcePDFScholar