← Search

I-Bin Liao

4 accepted papers

2025

Memory-Augmented Re-Completion for 3D Semantic Scene Completion

AAAI 2025technical

Semantic Scene Completion (SSC) aims to reconstruct a 3D voxel representation occupied by semantic classes based on ordinary inputs such as 2D RGB images, depth maps, or point clouds. Given the cost-effective and promising applications in autonomous driving, camera-based SSC has attracted considerab…

2025

PAVLM: Advancing Point Cloud based Affordance Understanding Via Vision-Language Model

IROS 2025

Affordance understanding, the task of identifying actionable regions on 3D objects, plays a vital role in allowing robotic systems to engage with and operate within the physical world. Although Visual Language Models (VLMs) have excelled in high-level reasoning and long-horizon planning for robotic

Cited by 6SourcecodeScholar
2023

Location-Aware Visual Question Generation with Lightweight Models

EMNLP 2023long main

This work introduces a novel task, location-aware visual question generation (LocaVQG), which aims to generate engaging questions from data relevant to a particular geographical location. Specifically, we represent such location-aware information with surrounding images and a GPS coordinate. To tack…

Cited by 0SourcecodeScholar
2016

Structural maximum a posteriori speaker adaptation of speaking rate-dependent hierarchical prosodic model for Mandarin TTS

ICASSP 2016accepted

In this paper, a structural maximum a posterior speaker adaptation method to adjust the existing speaking rate (SR) dependent hierarchical prosodic model (SR-HPM) to a new speaker's data for realizing a new voice of any given SR is discussed. The adaptive SR-HPM is formulated based on MAP estimation…

Cited by 0SourceScholar