← Search

Xue Zhang

25 accepted papers

2025

AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation

ACL 2025long

In modern large language models (LLMs), LLM alignment is of crucial importance and is typically achieved through methods such as reinforcement learning from human feedback (RLHF) and direct preference optimization (DPO). However, in most existing methods for LLM alignment, all tokens in the response…

2025

CM-Align: Consistency-based Multilingual Alignment for Large Language Models

EMNLP 2025

Current large language models (LLMs) generally show a significant performance gap in alignment between English and other languages.To bridge this gap, existing research typically leverages the model’s responses in English as a reference to select the best/worst responses in other languages, which ar

2025

Less, but Better: Efficient Multilingual Expansion for LLMs via Layer-wise Mixture-of-Experts

ACL 2025long

Continually expanding new languages for existing large language models (LLMs) is a promising yet challenging approach to building powerful multilingual LLMs.The biggest challenge is to make the model continuously learn new languages while preserving the proficient ability of old languages.To achieve…

2025

Multilingual Knowledge Editing with Language-Agnostic Factual Neurons

COLING 2025main

Multilingual knowledge editing (MKE) aims to simultaneously update factual knowledge across multiple languages within large language models (LLMs). Previous research indicates that the same knowledge across different languages within LLMs exhibits a degree of shareability. However, most existing MKE…

2025

SGDet3D: Semantics and Geometry Fusion for 3D Object Detection Using 4D Radar and Camera

RA-L 2025

4D millimeter-wave radar has gained attention as an emerging sensor for autonomous driving in recent years. However, existing 4D radar and camera fusion models often fail to fully exploit complementary information within each modality and lack deep cross-modal interactions. To address these issues,

Cited by 26SourceScholar
2024

Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects

ECCV 2024poster

"We interact with the world with our hands and see it through our own (egocentric) perspective. A holistic understanding of such interactions from egocentric views is important for tasks in robotics, AR/VR, action recognition and motion generation. Accurately reconstructing such interactions in is c…

2024

Disturbance Observer Based Terminal Sliding Mode Control for an Electromagnetic Actuated Micropositioner With Prescribed Performance

RA-L 2024

Precise and smooth control of an electromagnetic actuated micropositioner is a challenging task due to the system's nonlinear electromagnetic dynamics, as well as its unique structure and application scenarios. The main challenges include high-order nonlinearity, modeling uncertainties, input buffet

Cited by 4SourceScholar
2024

Dual-Space Knowledge Distillation for Large Language Models

EMNLP 2024main

Knowledge distillation (KD) is known as a promising solution to compress large language models (LLMs) via transferring their knowledge to smaller models. During this process, white-box KD methods usually minimize the distance between the output distributions of the two models so that more knowledge…

2024

From Fourier to Neural ODEs: Flow Matching for Modeling Complex Systems

ICML 2024poster

Modeling complex systems using standard neural ordinary differential equations (NODEs) often faces some essential challenges, including high computational costs and susceptibility to local optima. To address these challenges, we propose a simulation-free framework, called Fourier NODEs (FNODEs), tha…

Cited by 6SourcePDFScholar
2024

Graph-Enhanced Hybrid Sampling for Multi-Armed Bandit Recommendation

ICASSP 2024accepted

Graph-based multi-armed bandit algorithms utilize the relationship between users to select the best item to recommend for maximal reward, which is decided by items’ features and un-known users’ preferences. Therefore, the precise estimation of users’ preferences is fairly important and indispensable…

Cited by 0SourceScholar
2023

A Quality-based Syntactic Template Retriever for Syntactically-Controlled Paraphrase Generation

EMNLP 2023long main

Existing syntactically-controlled paraphrase generation (SPG) models perform promisingly with human-annotated or well-chosen syntactic templates. However, the difficulty of obtaining such templates actually hinders the practical application of SPG models. For one thing, the prohibitive cost makes it…

Cited by 0SourcecodeScholar
2023

An Autonomous Surgical Instrument Tracking Framework With a Binocular Camera for a Robotic Flexible Laparoscope

RA-L 2023

In minimally invasive surgery (MIS), the field of view (FOV) plays a vital role. To enhance the stability of FOV and lighten the burden on surgeons, robot-assisted laparoscope systems have been developed and introduced into surgery. However, most of the existing automatic surgical tool tracking sche

Cited by 13SourceScholar
2022

A Kinematic Modeling and Control Scheme for Different Robotic Endoscopes: A Rudimentary Research Prototype

RA-L 2022

In image-guided robotic surgery, there exist different endoscopes coupled either with specialized surgical robots (SSRs) or general industrial robots (GIRs). In general, SSRs mechanically respect the remote-center-of-motion (RCM) constraints with directly and explicitly controllable degrees-of-freed

Cited by 5SourceScholar
2022

A Surgeon Preference-Guided Autonomous Instrument Tracking Method With a Robotic Flexible Endoscope Based on dVRK Platform

RA-L 2022

In minimally invasive surgery, endoscopes serve as the eyes of surgeon. To avoid fatigue in manual endoscope steering, robotic endoscope holders have been developed. Unfortunately, existing robotic endoscope holders are not widely adopted due to the poor surgeon-robot cooperation. In this work, we d

Cited by 30SourceScholar
2022

Generating Authentic Adversarial Examples beyond Meaning-preserving with Doubly Round-trip Translation

NAACL 2022long

Generating adversarial examples for Neural Machine Translation (NMT) with single Round-Trip Translation (RTT) has achieved promising results by releasing the meaning-preserving restriction. However, a potential pitfall for this approach is that we cannot decide whether the generated examples are adv…

2021

An Autonomous Robotic Flexible Endoscope System with a DNA-inspired Continuum Mechanism

ICRA 2021poster

In this paper, we proposed an autonomous robotic flexible endoscope system for the laparoscopic bariatric surgery (LBS). This system comprises a UR5 robot and a flexible endoscope equipped with a novel continuum joint, named reinforced double helix continuum mechanism. Compared with the simple helix…

Cited by 14SourceScholar
2021

Orientation Control of an Electromagnetically Actuated Soft-Tethered Colonoscope Based on 2OR Pseudo-Rigid-Body Model

ICRA 2021poster

Colorectal cancer incidence has been steadily rising worldwide. Magnetic colonoscopes provide new approaches to conduct colon inspection and treatment. This paper presents a novel electromagnetically actuated soft-tethered colonoscope to achieve precise and stable orientation control. An inflated ba…

Cited by 9SourceScholar
2020

A Novel Flexible Robotic Endoscope With Constrained Tendon-Driven Continuum Mechanism

RA-L 2020

This letter presents a novel flexible robotic endoscope with a constrained tendon-driven continuum mechanism (CTCM) which is targeted for bariatric surgery. The robotic endoscope is composed of a UR5 robot and a CTCM-based flexible endoscope. By introducing a constraint tube, both the angulation and

Cited by 46SourceScholar
2020

Sparse Directed Graph Learning for Head Movement Prediction in 360 Video Streaming

ICASSP 2020accepted

High-definition 360 videos encoded in fine quality are typically too large in size to stream in its entirety over bandwidth (BW)-constrained networks. One popular remedy is to interactively extract and send a spatial sub-region corresponding to a viewer's current field-of-view (FoV) in a head-mounte…

Cited by 0SourceScholar
2018

A Biomimetic Soft Robot for Inspecting Pipeline with Significant Diameter Variation

IROS 2018poster

Navigation through tubular environment is fundamental in tasks such as pipeline inspection, gastrointestinal tract inspection, etc. Conventional pipeline inspection robots are mostly made by rigid materials and could not well adapt to the large size variation of the environment. Soft robots provide…

Cited by 49SourceScholar
2018

A Novel Magnetic Anchored and Steered Camera Robot for Single Port Access Surgery

ICRA 2018poster

This paper presents a novel magnetic anchored and steered camera robot intended for minimally invasive surgery (MIS), particularly for single port access (SPA) surgery. The design aims to achieve both compactness and a planar pan/tilt workspace (instead of hemispheric) to lower robot footprint in ve…

Cited by 10SourceScholar
2017

Mobile phone clustering from acquired speech recordings using deep Gaussian supervector and spectral clustering

ICASSP 2017accepted

Acquisition device clustering from speech recordings is a new and critical problem in the field of speech forensic, which aims at merging speech recordings acquired by the same device into one cluster without both pre-knowing prior information of the processed data and pre-training classifier. We pr…

Cited by 0SourceScholar