← Search

Tao Gao

14 accepted papers

2025

Brain-Inspired Spiking Neural Networks for Energy-Efficient Object Detection

CVPR 2025poster

Brain-inspired spiking neural networks (SNNs) have the capability of energy-efficient processing of temporal information. However, leveraging the rich dynamic characteristics of SNNs and prior works in artificial neural networks (ANNs) to construct an effective object detection model for visual task…

Cited by 0SourcePDFScholar
2025

Inverse Attention Agents for Multi-Agent Systems

ICLR 2025poster

A major challenge for Multi-Agent Systems (MAS) is enabling agents to adapt dynamically to diverse environments in which opponents and teammates may continually change. Agents trained using conventional methods tend to excel only within the confines of their training cohorts; their performance drops…

2025

Model-Based Control Strategies Comparison of One Bionic Ankle Tensegrity Exoskeleton: BATE

ICRA 2025

This paper presents a comparative analysis of model-based control strategies for a Bionic Ankle Tensegrity Exoskeleton (BATE), designed to emulate the self-stress equilibrium and self-supporting characteristics of the human ankle biotensegrity structure. Model-based control strategies are convention

Cited by 1SourceScholar
2025

Multi-axis Prompt and Multi-dimension Fusion Network for All-in-one Weather-degraded Image Restoration

AAAI 2025technical

Existing approaches aiming to remove adverse weather degradations compromise the image quality and incur the long processing time. To this end, we introduce a multi-axis prompt and multi-dimension fusion network (MPMF-Net). Specifically, we develop a multi-axis prompts learning block (MPLB), which l…

2024

Encoder-Minimal and Decoder-Minimal Framework for Remote Sensing Image Dehazing

ICASSP 2024accepted

Haze obscures remote sensing images, hindering valuable information extraction. To this end, we propose RSHazeNet, an encoder-minimal and decoder-minimal framework for efficient remote sensing image dehazing. Specifically, regarding the process of merging features within the same level, we develop a…

Cited by 0SourceScholar
2024

Kinematic Modeling of Twisted String Actuator Based on Invertible Neural Networks

IROS 2024poster

Twisted String Actuators (TSAs) exhibit several advantages, including lightweight, compact, and having a high power-to-weight ratio. However, current research on kinematic models of TSAs is limited to deriving the relationship between motor input and output through idealized geometric calculations.…

Cited by 0SourceScholar
2024

Multi-Dimension Queried and Interacting Network for Stereo Image Deraining

ICASSP 2024accepted

Eliminating the rain degradation in stereo images poses a formidable challenge, which necessitates the efficient exploitation of mutual information present between the dual views. To this end, we devise MQINet, which employs multi-dimension queries and interactions for stereo image deraining. More s…

Cited by 0SourceScholar
2024

Novel Multiport Output Twisted String Actuator with Self-differential Mechanism: Hand Glove Application

IROS 2024poster

The differential mechanism can reduce the number of actuators and efficiently distribute force or power. We proposed a novel multiport output twisted string actuator (MO-TSA) with self-differential mechanism that employs a single actuator to achieve multiport outputs. The differential MO-TSA is adap…

Cited by 0SourceScholar
2023

LRRU: Long-short Range Recurrent Updating Networks for Depth Completion

ICCV 2023poster

Existing deep learning-based depth completion methods generally employ massive stacked layers to predict the dense depth map from sparse input data. Although such approaches greatly advance this task, their accompanied huge computational complexity hinders their practical applications. To accomplish…

Cited by 55PDFcodeScholar
2022

Emergent Graphical Conventions in a Visual Communication Game

NeurIPS 2022accept

Humans communicate with graphical sketches apart from symbolic languages. Primarily focusing on the latter, recent studies of emergent communication overlook the sketches; they do not account for the evolution process through which symbolic sign systems emerge in the trade-off between iconicity and…

Cited by 19SourcePDFScholar
2021

Learning Triadic Belief Dynamics in Nonverbal Communication From Videos

CVPR 2021poster

Humans possess a unique social cognition capability; nonverbal communication can convey rich social information among agents. In contrast, such crucial social characteristics are mostly missing in the existing scene understanding literature. In this paper, we incorporate different nonverbal communic…

Cited by 27PDFcodeScholar
2021

YouRefIt: Embodied Reference Understanding With Language and Gesture

ICCV 2021poster

We study the machine's understanding of embodied reference: One agent uses both language and gesture to refer to an object to another agent in a shared physical environment. Of note, this new visual task requires understanding multimodal cues with perspective-taking to identify which object is being…

Cited by 46PDFScholar
2020

Joint Inference of States, Robot Knowledge, and Human (False-)Beliefs

ICRA 2020poster

Aiming to understand how human (false-)belief— a core socio-cognitive ability—would affect human interactions with robots, this paper proposes to adopt a graphical model to unify the representation of object states, robot knowledge, and human (false-)beliefs. Specifically, a parse graph (pg) is lear…

Cited by 27SourceScholar
2016

Inferring human intent from video by sampling hierarchical plans

IROS 2016poster

This paper presents a method which allows robots to infer a human's hierarchical intent from partially observed RGBD videos by imagining how the human will behave in the future. This capability is critical for creating robots which can interact socially or collaboratively with humans. We represent i…

Cited by 43SourceScholar