← Search

Ziwei Deng

5 accepted papers

2025

DoF: A Diffusion Factorization Framework for Offline Multi-Agent Reinforcement Learning

ICLR 2025poster

Diffusion models have been widely adopted in image and language generation and are now being applied to reinforcement learning. However, the application of diffusion models in offline cooperative Multi-Agent Reinforcement Learning (MARL) remains limited. Although existing studies explore this direct…

2025

PlanU: Large Language Model Reasoning through Planning under Uncertainty

NeurIPS 2025poster

Large Language Models (LLMs) are increasingly being explored across a range of reasoning tasks. However, LLMs sometimes struggle with reasoning tasks under uncertainty that are relatively easy for humans, such as planning actions in stochastic environments. The adoption of LLMs for reasoning is impe…

Cited by 0SourceScholar
2024

Melting Pot Contest: Charting the Future of Generalized Cooperative Intelligence

NeurIPS 2024poster

Multi-agent AI research promises a path to develop human-like and human-compatible intelligent technologies that complement the solipsistic view of other approaches, which mostly do not consider interactions between agents. Aiming to make progress in this direction, the Melting Pot contest 2023 focu…

Cited by 0SourcePDFScholar
2020

Cycle-Contrast for Self-Supervised Video Representation Learning

NeurIPS 2020poster

We present Cycle-Contrastive Learning (CCL), a novel self-supervised method for learning video representation. Following a nature that there is a belong and inclusion relation of video and its frames, CCL is designed to find correspondences across frames and videos considering the contrastive repres…

Cited by 54SourcePDFScholar
2019

MMAct: A Large-Scale Dataset for Cross Modal Human Action Understanding

ICCV 2019poster

Unlike vision modalities, body-worn sensors or passive sensing can avoid the failure of action understanding in vision related challenges, e.g. occlusion and appearance variation. However, a standard large-scale dataset does not exist, in which different types of modalities across vision and sensors…

Cited by 124PDFScholar