← Search

Ziluo Ding

14 accepted papers

2025

CROSSER: Learning Generalizable Humanoid Locomotion Through Inverse Dynamics-Guided Cross-Simulator Adaptation

RA-L 2025

The reality gap between simulation and real-world dynamics critically hinders the deployment of robust humanoid locomotion policies, as policies trained in a single simulator often overfit to domain-specific dynamics. To address this challenge, we propose CROSSER (Inverse Dynamics-Guided Cross-Simul

Cited by 1SourceScholar
2025

Cradle: Empowering Foundation Agents towards General Computer Control

ICML 2025poster

Despite their success in specific scenarios, existing foundation agents still struggle to generalize across various virtual scenarios, mainly due to the dramatically different encapsulations of environments with manually designed observation and action spaces. To handle this issue, we propose the Ge…

2025

From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots

NeurIPS 2025spotlight

Achieving general agile whole-body control on humanoid robots remains a major challenge due to diverse motion demands and data conflicts. While existing frameworks excel in training single motion-specific policies, they struggle to generalize across highly varied behaviors due to conflicting control…

Cited by 0SourceScholar
2024

Learning to Robustly Reconstruct Dynamic Scenes from Low-light Spike Streams

ECCV 2024poster

"Spike camera with high temporal resolution can fire continuous binary spike streams to record per-pixel light intensity. By using reconstruction methods, the scene details in high-speed scenes can be restored from spike streams. However, existing methods struggle to perform well in low-light enviro…

2024

Multi-Agent Coordination via Multi-Level Communication

NeurIPS 2024poster

The partial observability and stochasticity in multi-agent settings can be mitigated by accessing more information about others via communication. However, the coordination problem still exists since agents cannot communicate actual actions with each other at the same time due to the circular depend…

Cited by 0SourcePDFScholar
2024

Reinforcement Learning Friendly Vision-Language Model for Minecraft

ECCV 2024poster

"One of the essential missions in the AI research community is to build an autonomous embodied agent that can achieve high-level performance across a wide spectrum of tasks. However, acquiring or manually designing rewards for all open-ended tasks is unrealistic. In this paper, we propose a novel cr…

2024

Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing

AAAI 2024technical

Exploration in decentralized cooperative multi-agent reinforcement learning faces two challenges. One is that the novelty of global states is unavailable, while the novelty of local observations is biased. The other is how agents can explore in a coordinated way. To address these challenges, we prop…

2023

Entity Divider with Language Grounding in Multi-Agent Reinforcement Learning

ICML 2023poster

We investigate the use of natural language to drive the generalization of policies in multi-agent settings. Unlike single-agent settings, the generalization of policies should also consider the influence of other agents. Besides, with the increasing number of entities in multi-agent settings, more a…

2023

Unsupervised Optical Flow Estimation with Dynamic Timing Representation for Spike Camera

NeurIPS 2023poster

Efficiently selecting an appropriate spike stream data length to extract precise information is the key to the spike vision tasks. To address this issue, we propose a dynamic timing representation for spike streams. Based on multi-layers architecture, it applies dilated convolutions on temporal dime…

2022

Modeling The Detection Capability Of High-Speed Spiking Cameras

ICASSP 2022accepted

The novel working principle enables spiking cameras to capture high-speed moving objects. However, the applications of spiking cameras can be affected by many factors, such as brightness intensity, detectable distance, and the maximum speed of moving targets. Improper settings such as weak ambient b…

Cited by 0SourceScholar
2022

Spatio-Temporal Recurrent Networks for Event-Based Optical Flow Estimation

AAAI 2022technical

Event camera has offered promising alternative for visual perception, especially in high speed and high dynamic range scenes. Recently, many deep learning methods have shown great success in providing model-free solutions to many event-based problems, such as optical flow estimation. However, existi…

2020

Learning Individually Inferred Communication for Multi-Agent Cooperation

NeurIPS 2020oral

Communication lays the foundation for human cooperation. It is also crucial for multi-agent cooperation. However, existing work focuses on broadcast communication, which is not only impractical but also leads to information redundancy that could even impair the learning process. To tackle these diff…