← Search

Hongbo Wang

21 accepted papers

2026

BOOSTAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models

ICML 2026poster

Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually fix bugs. We present BoostAPR, a three-stage framework: (1) supervised fine-tuning on execution-verified demonstrations with reasoning traces, (2) tr…

Cited by 0SourceScholar
2026

CoGrad3D: Spatially-Coupled Timestep Optimization with Orthogonal Gradient Fusion for 3D Generation

AAAI 2026technical

Score Distillation Sampling has driven recent advances in text-to-3D generation. However, current approaches often fail to produce 3D assets that are both rich in detail and consistent across viewpoints. These limitations primarily arise from imbalanced guidance on fine-grained details and an overde

Cited by 0SourcePDFScholar
2026

Coloring the Noise: Adversarial Sobolev Alignment for Faithful Image Super Resolution

ICML 2026poster

Generative priors in Image Super-Resolution (SR) often compromise faithful restoration, we attribute this limitation to a fundamental spectral misalignment between isotropic objectives and the intrinsic natural image manifold. While Direct Preference Optimization offers a path to alignment, its reli…

Cited by 0SourceScholar
2026

DRAFT-RL: Multi-Agent Chain-of-Draft Reasoning for Reinforcement Learning-Enhanced LLMs

AAAI 2026technical

Large Language Models (LLMs) have shown impressive capabilities in multi-step reasoning and problem-solving. Recent works introduce multi-agent reflection frameworks where multiple LLM agents critique and refine each other’s outputs using reinforcement learning (RL). However, these approaches often

Cited by 0SourcePDFScholar
2026

Failure Detection With Zero-Shot Error Correction in Robotic Manipulation

RA-L 2026

Diffusion Policy (DP) is effective for imitation learning in robotic manipulation, yet likelihood-based replanning methods lack self-correction and often fail once execution deviates. Vision-Language Models (VLMs) offer strong spatial reasoning for failure handling but incur prohibitive latency for

Cited by 0SourceScholar
2026

PROPAGATING SIMILARITY, MITIGATING UNCERTAINTY: SIMILARITY PROPAGATION-ENHANCED UNCERTAINTY FOR MULTIMODAL RECOMMENDATION

ICASSP 2026poster

Multimodal Recommendation (MMR) systems are crucial for modern platforms but are often hampered by inherent noise and uncertainty in modal features, such as blurry images, diverse visual appearances, or ambiguous text. Existing methods often overlook this modality-specific uncertainty, leading to in…

Cited by 0SourcePDFScholar
2026

Think-Then-Generate: Structural Chain-of-Thought Reasoning for Consistent 3D Generation

CVPR 2026

Recently, generating 3D assets using visual priors from pretrained diffusion models has shown remarkable results. However, due to the inherent lack of 3D geometric priors in 2D diffusion, the synthesized results often suffer from spatial hallucination and multi-view inconsistency. To address this li

Cited by 0SourcecodeScholar
2026

Towards Fine-Grained Attribution: Instance-Aware Preference Optimization for Aligning Diffusion Models

CVPR 2026

Direct Preference Optimization has achieved remarkable success in aligning diffusion models with human feedback. However, existing methods heavily rely on image-level preferences, which suffer from sparse rewards in the spatial dimension. This creates a fundamental misalignment: while an image may b

Cited by 0SourceScholar
2025

Design and Validation of a Non-Contact Bed Micro-Movement Sensing System for HRV Monitoring

RA-L 2025

The growing demand for non-wearable, long-term health monitoring solutions, especially for elderly, has driven the development of bed-based physiological monitoring systems. This letter introduces a novel non-contact, bed-based micro-movement sensing system designed for accurate heart rate variabili

Cited by 0SourceScholar
2025

Instance Segmentation of Airway Anatomies Using Mask R-CNN Prompt Adaptation-SAM

ICASSP 2025accepted

Accurate identification of key anatomy structures in airway intubation, the primary step in general anesthesia, is crucial for surgical success and patient safety. Achieving both object detection and segmentation in this context using deep learning technologies is challenging due to limited labeled…

Cited by 0SourceScholar
2025

Microbubble Sheath Empowered Pneumatic Artificial Muscles for Highly-Precise and Stable Needle Insertion

RA-L 2025

Soft pneumatic actuators and robotic systems offer significant advantages in biomedical applications and enable tasks beyond the capabilities of rigid systems, benefitting from their inherent deformability, compliance, and adaptability. However, their low stiffness often leads to severe vibrations d

Cited by 0SourceScholar
2025

Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Detector

ICASSP 2025accepted

Patronizing and Condescending Language (PCL) is a form of discriminatory toxic speech targeting vulnerable groups, threatening both online and offline safety. While toxic speech research has mainly focused on overt toxicity, such as hate speech, microaggressions in the form of PCL remain underexplor…

Cited by 0SourceScholar
2024

A Perceptive Pneumatic Artificial Muscle Empowered by Double Helix Fiber Reinforcement

IROS 2024poster

In the last decades, soft robotics has been growing rapidly as an emerging research topic, bringing new paradigms for robotic manipulation, locomotion, and human‒machine interactions. Pneumatic artificial muscle is a powerful, lightweight, rapid response with great design flexibility, making it prom…

Cited by 0SourceScholar
2024

Hallo3D: Multi-Modal Hallucination Detection and Mitigation for Consistent 3D Content Generation

NeurIPS 2024poster

Recent advancements in 3D content generation have been significant, primarily due to the visual priors provided by pretrained diffusion models. However, large 2D visual models exhibit spatial perception hallucinations, leading to multi-view inconsistency in 3D content generated through Score Distill…

Cited by 1SourcePDFScholar
2024

PclGPT: A Large Language Model for Patronizing and Condescending Language Detection

EMNLP 2024finding

Disclaimer: Samples in this paper may be harmful and cause discomfort! Patronizing and condescending language (PCL) is a form of speech directed at vulnerable groups. As an essential branch of toxic language, this type of language exacerbates conflicts and confrontations among Internet communities a…

2023

Compact Waist Rehabilitation Robot Inspired by McKenzie Therapy: Design, Analysis and Validation

RA-L 2023

Aiming at solving the problems of low compliance and unsatisfactory therapeutic effect in most existing waist rehabilitation robots, this letter presents a compact waist rehabilitation robot (CWRR) for patients with low back pain (LBP). First, the CWRR with a rigid pose adjustment mechanism and a so

Cited by 2SourceScholar
2023

DocRED-FE: A Document-Level Fine-Grained Entity and Relation Extraction Dataset

ICASSP 2023accepted

Joint entity and relation extraction (JERE) is one of the most important tasks in information extraction. However, most existing works focus on sentence-level coarse-grained JERE, which have limitations in real-world scenarios. In this paper, we construct a large-scale document-level fine-grained JE…

Cited by 0SourceScholar
2023

FABRIKv: A Fast, Iterative Inverse Kinematics Solver for Surgical Continuum Robot with Variable Curvature Model

IROS 2023poster

Due to the advantages of high flexibility, large workspace, and good human-body compatibility, flexible tendon-driven surgical continuum robots have attracted a lot of attention in robot-assisted minimally invasive surgery. However, due to the coupling of the position and angle of the continuum robo…

Cited by 1SourceScholar
2022

FIORA : A Flexible Tendon-Driven Continuum Manipulator for Laparoscopic Surgery

RA-L 2022

Aiming at solving the problems of low flexibility, poor internal extension, and chopstick effect in most laparoscopic surgical robots, this letter presents a flexible tendon-driven continuum surgical manipulator with eight degrees of freedom, called FIORA. The manipulator is composed of three indepe

Cited by 29SourceScholar
2017

A soft multi-axial force sensor to assess tissue properties in RealTime

IROS 2017poster

Objective: This work presents a method for the use of a soft multi-axis force sensor to determine tissue trauma in Minimally Invasive Surgery. Despite recent developments, there is a lack of effective haptic sensing technology employed in instruments for Minimally Invasive Surgery (MIS). There is th…

Cited by 7SourceScholar
2015

A model predictive control approach for the Partner Ballroom Dance Robot

ICRA 2015poster

A model predictive controller is developed for following the position of a human dancer in robot ballroom dancing. The control design uses a dynamic model of a dancer, based on a variant of the so-called 3D Linear Inverted Pendulum Mode that includes also the swing foot. This model serves as a basis…

Cited by 7SourceScholar