← Search

Haichuan Gao

8 accepted papers

2026

DeFacto: Counterfactual Thinking with Images for Enforcing Evidence-Grounded and Faithful Reasoning

ICML 2026poster

Recent advances in multimodal language models (MLLMs) have made thinking with images a dominant paradigm for multimodal reasoning. However, existing methods still fail to ensure evidence–answer consistency, where correct answers must be supported by correct visual evidence. To address this issue, we…

Cited by 0SourceScholar
2025

Adaptive Fission: Post-training Encoding for Low-latency Spike Neural Networks

NeurIPS 2025poster

Spiking Neural Networks (SNNs) often rely on rate coding, where high-precision inference depends on long time-steps, leading to significant latency and energy cost—especially for ANN-to-SNN conversions. To address this, we propose Adaptive Fission, a post-training encoding technique that selectively…

Cited by 0SourceScholar
2025

OURO: A Self-Bootstrapped Framework for Enhancing Multimodal Scene Understanding

ICCV 2025poster

Multimodal large models have made significant progress, yet fine-grained understanding of complex scenes remains a challenge. High-quality, large-scale vision-language datasets are essential for addressing this issue. However, existing methods often rely on labor-intensive manual annotations or clos…

2024

Spatio-Temporal Approximation: A Training-Free SNN Conversion for Transformers

ICLR 2024poster

Spiking neural networks (SNNs) are energy-efficient and hold great potential for large-scale inference. Since training SNNs from scratch is costly and has limited performance, converting pretrained artificial neural networks (ANNs) to SNNs is an attractive approach that retains robust performance wi…

Cited by 12SourcePDFScholar
2023

Fast Counterfactual Inference for History-Based Reinforcement Learning

AAAI 2023technical

Incorporating sequence-to-sequence models into history-based Reinforcement Learning (RL) provides a general way to extend RL to partially-observable tasks. This method compresses history spaces according to the correlations between historical observations and the rewards. However, they do not adjust…

Cited by 3SourcePDFScholar
2021

CRIL: Continual Robot Imitation Learning via Generative and Prediction Model

IROS 2021poster

Imitation learning (IL) algorithms have shown promising results for robots to learn skills from expert demonstrations. However, they need multi-task demonstrations to be provided at once for acquiring diverse skills, which is difficult in real world. In this work we study how to realize continual im…

Cited by 20SourcecodeScholar
2020

Adaptability Preserving Domain Decomposition for Stabilizing Sim2Real Reinforcement Learning

IROS 2020poster

In sim-to-real transfer of Reinforcement Learning (RL) policies for robot tasks, Domain Randomization (DR) is a widely used technique for improving adaptability. However, in DR there is a conflict between adaptability and training stability, and heavy DR tends to result in instability or even failur…

Cited by 6SourceScholar