← Search

Zhi Liu

24 accepted papers

2026

Biarticular Rigid Powered Lower Extremity Exoskeleton Robot

ICRA 2026poster

Lower extremity exoskeletons designed for multi-joint assistance are increasingly explored for rehabilitation and human augmentation. However, conventional monoarticular designs often suffer from joint misalignment and actuator redundancy, limiting their efficiency and user comfort. This study prese…

Cited by 0SourceScholar
2026

M4-SAM: Multi-Modal Mixture-of-Experts with Memory-Augmented SAM for RGB-D Video Salient Object Detection

CVPR 2026

The Segment Anything Model 2 (SAM2) has emerged as a foundation model for universal segmentation. Owing to its generalizable visual representations, SAM2 has been successfully applied to various downstream tasks. However, extending SAM2 to the RGB-D video salient object detection (RGB-D VSOD) task e

Cited by 0SourcecodeScholar
2026

SAM-DAQ: Segment Anything Model with Depth-guided Adaptive Queries for RGB-D Video Salient Object Detection

AAAI 2026technical

Recently segment anything model (SAM) has attracted widespread concerns, and it is often treated as a vision foundation model for universal segmentation. Some researchers have attempted to directly apply the foundation model to the RGB-D video salient object detection (RGB-D VSOD) task, which often

Cited by 0SourcePDFScholar
2026

SEMC: Structure-Enhanced Mixture-of-Experts Contrastive Learning for Ultrasound Standard Plane Recognition

AAAI 2026technical

Ultrasound standard plane recognition is essential for clinical tasks such as disease screening, organ evaluation, and biometric measurement. However, existing methods fail to effectively exploit shallow structural information and struggle to capture fine-grained semantic differences through contras

Cited by 0SourcePDFScholar
2026

SMixer: Rethinking Efficient-Training and Event-Driven SNNs

ICLR 2026poster

Spiking Neural Networks (SNNs) offer a promising, energy-efficient paradigm for computation, but their practical application is hindered by challenges in architecture design and training costs. For example, Spiking ResNet exhibits relatively low performance, whereas high-performance Spiking Transfor…

Cited by 0SourceScholar
2025

C2PD: Continuity-Constrained Pixelwise Deformation for Guided Depth Super-Resolution

AAAI 2025technical

Guided depth super-resolution (GDSR) has demonstrated impressive performance across a wide range of domains, with numerous methods being proposed. However, existing methods often treat depth maps as images, where shading values are computed discretely, making them struggle to effectively restore the…

2025

KELE: A Multi-Agent Framework for Structured Socratic Teaching with Large Language Models

EMNLP 2025

Socratic teaching, known for its emphasis on heuristic questioning and deep thinking, has demonstrated significant advantages in promoting students’ cognitive development. However, traditional Socratic teaching places high demands on teachers’ expertise and real-time feedback capabilities, making it

2025

LLM-Driven Implicit Target Augmentation and Fine-Grained Contextual Modeling for Zero-Shot and Few-Shot Stance Detection

EMNLP 2025

Stance detection aims to identify the attitude expressed in text towards a specific target. Recent studies on zero-shot and few-shot stance detection focus primarily on learning generalized representations from explicit targets. However, these methods often neglect implicit yet semantically importan

2025

MESC-3D:Mining Effective Semantic Cues for 3D Reconstruction from a Single Image

CVPR 2025poster

Reconstructing 3D shapes from a single image plays an important role in computer vision. Many methods have been proposed and achieve impressive performance. However, existing methods mainly focus on extracting semantic information from images and then simply concatenating it with 3D point clouds wit…

2025

SGTC: Semantic-Guided Triplet Co-training for Sparsely Annotated Semi-Supervised Medical Image Segmentation

AAAI 2025technical

Although semi-supervised learning has made significant advances in the field of medical image segmentation, fully annotating a volumetric sample slice by slice remains a costly and time-consuming task. Even worse, most of the existing approaches pay much attention to image-level information and igno…

2025

The Anti-Misalignment Mechanism of Bionic Knee Joint of Lower Limb Exoskeleton Based on Spherical Cross Four-Bar

IROS 2025

To minimize discomfort and injury risk in exoskeleton users, this paper addresses the misalignment between the device and the human knee joint. The knee's spatial motion complexity, characterized by multi-planar rotation axes as flexion angle changes, cannot be accurately replicated by existing sing

Cited by 0SourceScholar
2024

Identifying and Addressing Disparities in Public Libraries with Bayesian Latent Variable Modeling

AAAI 2024technical

Public libraries are an essential public good. We ask: are urban library systems providing equitable service to all residents, in terms of the books they have access to and check out? If not, what causes disparities: heterogeneous book collections, resident behavior and access, and/or operational po…

2024

The Control Strategy for Vehicle Transfer Robots in RO/RO Terminal Environments

IROS 2024poster

In the labor-intensive Roll-On/Roll-Off (RO/RO) terminal environment, research on vehicle transport robots with mobility, stability, and reliability is receiving increasing attention. This paper presents a novel control framework for a Straddle-Type Dual-Body vehicle transfer robot. Initially, fine…

Cited by 0SourceScholar
2023

Towards Reliable Item Sampling for Recommendation Evaluation

AAAI 2023technical

Since Rendle and Krichene argued that commonly used sampling-based evaluation metrics are ``inconsistent'' with respect to the global metrics (even in expectation), there have been a few studies on the sampling-based recommender system evaluation. Existing methods try either mapping the sampling-bas…

Cited by 12SourcePDFScholar
2022

Vision-based Uneven BEV Representation Learning with Polar Rasterization and Surface Estimation

CoRL 2022poster

In this work, we propose PolarBEV for vision-based uneven BEV representation learning. To adapt to the foreshortening effect of camera imaging, we rasterize the BEV space both angularly and radially, and introduce polar embedding decomposition to model the associations among polar grids. Polar gri…

Cited by 26SourcecodeScholar
2021

Climbot-Ω: A Soft Robot with Novel Grippers and Rigid-compliantly Constrained Body for Climbing on Various Poles

IROS 2021poster

Soft climbing robots have been attracting increasing attention in soft robotics community, and a lot of prototypes been proposed with basic climbing function implemented. Climbing on poles is a challenge with soft robots, and the capability of current pole-climbing soft robots needs to be improved i…

Cited by 3SourceScholar
2021

On Estimating Recommendation Evaluation Metrics under Sampling

AAAI 2021technical

Since the recent studies (KDD'20) done by Krichene and Rendle on the sampling based top-k evaluation metric for recommendation, there have been a lot of debate on the validity of using sampling for evaluating recommendation algorithms. Though their work and the recent work done by Li et. al. (KDD'…

Cited by 16SourcePDFScholar
2020

Cross-Modal Weighting Network for RGB-D Salient Object Detection

ECCV 2020poster

Depth maps contain geometric clues for assisting Salient Object Detection (SOD). In this paper, we propose a novel Cross-Modal Weighting (CMW) strategy to encourage comprehensive interactions between RGB and depth channels for RGB-D SOD. Specifically, three RGB-depth interaction modules, named CMW-L…

2019

Cleaning Adversarial Perturbations via Residual Generative Network for Face Verification

ICASSP 2019accepted

Deep neural networks (DNNs) have recently achieved impressive performances on various applications. However, recent researches show that DNNs are vulnerable to adversarial perturbations injected into input samples. In this paper, we investigate a defense method for face verification: a deep residual…

Cited by 0SourceScholar