← Search

Ziang Song

5 accepted papers

2022

Efficient Phi-Regret Minimization in Extensive-Form Games via Online Mirror Descent

NeurIPS 2022accept

A conceptually appealing approach for learning Extensive-Form Games (EFGs) is to convert them to Normal-Form Games (NFGs). This approach enables us to directly translate state-of-the-art techniques and analyses in NFGs to learning EFGs, but typically suffers from computational intractability due to…

Cited by 25SourcePDFScholar
2022

Learn To Remember: Transformer with Recurrent Memory for Document-Level Machine Translation

NAACL 2022findings

The Transformer architecture has led to significant gains in machine translation. However, most studies focus on only sentence-level translation without considering the context dependency within documents, leading to the inadequacy of document-level coherence. Some recent research tried to mitigate…

Cited by 20SourcePDFScholar
2022

When Can We Learn General-Sum Markov Games with a Large Number of Players Sample-Efficiently?

ICLR 2022poster

Multi-agent reinforcement learning has made substantial empirical progresses in solving games with a large number of players. However, theoretically, the best known sample complexity for finding a Nash equilibrium in general-sum games scales exponentially in the number of players due to the size of…

Cited by 124SourcePDFScholar