← Search

Xiaotian Zhang

11 accepted papers

2026

Beyond N-grams: A Hierarchical Reward Learning Framework for Clinically-Aware Medical Report Generation

AAAI 2026technical

Automatic medical report generation can greatly reduce the workload of doctors, but it is often unreliable for real-world deployment. Current methods can write formally fluent sentences but may be factually flawed, introducing serious medical errors known as clinical hallucinations, which make them

Cited by 0SourcePDFScholar
2026

UniVBench: Towards Unified Evaluation for Video Foundation Models

CVPR 2026

Video foundation models aim to integrate video understanding, generation, editing, and instruction following within a single framework, making them a central direction for next-generation multimodal systems. However, existing evaluation benchmarks remain fragmented and limited in scope, as they each

Cited by 0SourcecodeScholar
2025

DiffPO: Diffusion-styled Preference Optimization for Inference Time Alignment of Large Language Models

ACL 2025long

Inference-time alignment provides an efficient alternative for aligning LLMs with humans. However, these approaches still face challenges, such as limited scalability due to policy-specific value functions and latency during the inference phase. In this paper, we propose a novel approach, Diffusion-…

2025

Exploring Intrinsic Normal Prototypes within a Single Image for Universal Anomaly Detection

CVPR 2025poster

Anomaly detection (AD) is essential for industrial inspection, yet existing methods typically rely on "comparing" test images to normal references from a training set. However, variations in appearance and positioning often complicate the alignment of these references with the test image, limiting d…

2025

PAD: Personalized Alignment of LLMs at Decoding-time

ICLR 2025poster

Aligning with personalized preferences, which vary significantly across cultural, educational, and political differences, poses a significant challenge due to the computational costs and data demands of traditional alignment methods. In response, this paper presents Personalized Alignment at Decodin…

Cited by 10SourcePDFScholar
2025

Persona-judge: Personalized Alignment of Large Language Models via Token-level Self-judgment

ACL 2025finding

Aligning language models with human preferences presents significant challenges, particularly in achieving personalization without incurring excessive computational costs. Existing methods rely on reward signals and additional annotated data, limiting their scalability and adaptability to diverse hu…

Cited by 0SourcePDFScholar
2023

Investigating Glyph-Phonetic Information for Chinese Spell Checking: What Works and What’s Next?

ACL 2023findings

While pre-trained Chinese language models have demonstrated impressive performance on a wide range of NLP tasks, the Chinese Spell Checking (CSC) task remains a challenge. Previous research has explored using information such as glyphs and phonetics to improve the ability of CSC models to distinguis…

2023

Multijugate Dual Learning for Low-Resource Task-Oriented Dialogue System

ACL 2023findings

Dialogue data in real scenarios tend to be sparsely available, rendering data-starved end-to-end dialogue systems trained inadequately. We discover that data utilization efficiency in low-resource scenarios can be enhanced by mining alignment information uncertain utterance and deterministic dialogu…

Cited by 2SourcePDFScholar
2018

Augmented Joint Stiffness and Actuation Using Architectures of Soft Pneumatic Actuators

ICRA 2018poster

Soft robotic actuators are well suited for use in exoskeleton applications due to their innate compliance and low weight. We have developed a wearable soft robotic sleeve that uses fiber reinforced elastomeric enclosures (FREEs) to provide actuation and stiffness at the elbow for augmented lifting a…

Cited by 14SourceScholar