← Search

Zhihao Liu

13 accepted papers

2026

Autoregressive Meta-Actions for Unified Controllable Trajectory Generation in Autonomous Driving

RA-L 2026

Generating trajectories from high-level commands is critical for autonomous driving, but prevailing methods suffer from a flaw we term semantic misalignment. By associating long trajectories with a single, static meta-action (e.g., “lane change”), these methods corrupt training data during maneuver

Cited by 0SourcecodeScholar
2026

LandCraft: Designing the Structured 3D Landscapes via Text Guidance

AAAI 2026technical

Modeling large-scale landscapes is a foundational yet time-consuming task in many 3D applications, typically requiring substantial expertise. Recently, Text-to-3D techniques have emerged as a promising, beginner-friendly prototyping approach for generating 3D content from textual input. However, ex

Cited by 0SourcePDFScholar
2026

Peak-Return Greedy Slicing: Subtrajectory Selection for Transformer-based Offline RL

ICLR 2026poster

Offline reinforcement learning enables policy learning solely from fixed datasets, without costly or risky environment interactions, making it highly valuable for real-world applications. While Transformer-based approaches have recently demonstrated strong sequence modeling capabilities, they typica…

Cited by 0SourceScholar
2026

RLux-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models

RSS 2026poster

Recent advances in vision-language-action (VLA) models have motivated the extension of their capabilities to embodied settings, where reinforcement learning (RL) offers a principled way to optimize task success through interaction. However, existing methods remain fragmented, lacking both a unified …

Cited by 0SourceScholar
2025

BurstDeflicker: A Benchmark Dataset for Flicker Removal in Dynamic Scenes

NeurIPS 2025poster

Flicker artifacts in short-exposure images are caused by the interplay between the row-wise exposure mechanism of rolling shutter cameras and the temporal intensity variations of alternating current (AC)-powered lighting. These artifacts typically appear as uneven brightness distribution across the…

Cited by 0SourceScholar
2025

DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response

NeurIPS 2025poster

Large vision-language models (VLMs) have made great achievements in Earth vision. However, complex disaster scenes with diverse disaster types, geographic regions, and satellite sensors have posed new challenges for VLM applications. To fill this gap, we curate the first remote sensing vision-langua…

Cited by 0SourcecodeScholar
2025

FlareX: A Physics-Informed Dataset for Lens Flare Removal via 2D Synthesis and 3D Rendering

NeurIPS 2025poster

Lens flare occurs when shooting towards strong light sources, significantly degrading the visual quality of images. Due to the difficulty in capturing flare-corrupted and flare-free image pairs in the real world, existing datasets are typically synthesized in 2D by overlaying artificial flare templa…

Cited by 0SourceScholar
2024

Position: Rethinking Post-Hoc Search-Based Neural Approaches for Solving Large-Scale Traveling Salesman Problems

ICML 2024oral

Recent advancements in solving large-scale traveling salesman problems (TSP) utilize the heatmap-guided Monte Carlo tree search (MCTS) paradigm, where machine learning (ML) models generate heatmaps, indicating the probability distribution of each edge being part of the optimal solution, to guide MCT…

2024

Relaxing the Limitations of the Optimal Reciprocal Collision Avoidance Algorithm for Mobile Robots in Crowds

RA-L 2024

The Optimal Reciprocal Collision Avoidance (ORCA) algorithm is widely used for modeling agents in collision avoidance scenarios. However, suffering from limitations such as the improper reciprocal assumption that each agent is supposed to take half the responsibility for collision avoidance, the per

Cited by 5SourceScholar
2024

SVDTree: Semantic Voxel Diffusion for Single Image Tree Reconstruction

CVPR 2024poster

Efficiently representing and reconstructing the 3D geometry of biological trees remains a challenging problem in computer vision and graphics. We propose a novel approach for generating realistic tree models from single-view photographs. We cast the 3D information inference problem to a semantic vox…

2023

Hierarchical Multi-Agent Reinforcement Learning with Intrinsic Reward Rectification

ICASSP 2023accepted

Hierarchical reinforcement learning (HRL) is a promising approach to solving long-term decision problems and complex tasks, as high-level policy can guide the training procedure of low-level policy with macro actions and intrinsic rewards. However, the amount that macro actions influence decision-ma…

Cited by 0SourceScholar