← Search

Zheng WEI

9 accepted papers

2026

Adaptive Token Refinement in Long-Tailed Large Vision-Language Models Fine-Tuning

ICML 2026poster

While large vision-language models (LVLMs) have shown remarkable adaptability to downstream applications, their fine-tuning process remains susceptible to bias under long-tailed data. Compared to zero-shot scenarios, fine-tuning LVLMs on imbalanced datasets often yields limited performance improveme…

Cited by 0SourceScholar
2026

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

ICML 2026poster

Search agents powered by Large Language Models have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement learning has emerged as a powerful paradigm for training these agents to perform complex, multi-step reasoning. However, prior RL-based methods often rely on sp…

Cited by 0SourceScholar
2026

SesaHand: Enhancing 3D Hand Reconstruction via Controllable Generation with Semantic and Structural Alignment

ICLR 2026poster

Recent studies on 3D hand reconstruction have demonstrated the effectiveness of synthetic training data to improve estimation performance. However, most methods rely on game engines to synthesize hand images, which often lack diversity in textures and environments, and fail to include crucial compon…

Cited by 0SourceScholar
2025

ContextAware: A Multi-Agent Framework for Detecting Harmful Image-Based Comments on Social Media

IJCAI 2025

Detecting hidden stigmatization in social media poses significant challenges due to semantic misalignments between textual and visual modalities, as well as the subtlety of implicit stigmatization. Traditional approaches often fail to capture these complexities in real-world, multimodal content. To

2025

LIST: Linearly Incremental SQL Translator for Single-Hop Reasoning, Generation and Verification

ACL 2025finding

SQL languages often feature nested structures that require robust interaction with databases. Aside from the well-validated schema linking methods on PLMs and LLMs, we introduce the Linearly Incremental SQL Translator (LIST), a novel algorithmic toolkit designed to leverage the notable reasoning and…

2025

Uncertainty-Aware Iterative Preference Optimization for Enhanced LLM Reasoning

ACL 2025long

Direct Preference Optimization (DPO) has recently emerged as an efficient and effective method for aligning large language models with human preferences. However, constructing high-quality preference datasets remains challenging, often necessitating expensive manual or powerful LM annotations. Addit…

2024

SAM: A Self-Adaptive Attention Module for Context-Aware Recommendation System

ICASSP 2024accepted

Recently, textual information has been proven to positively affect recommendation systems. However, most of the existing methods only focus on representation learning of textual information in ratings, while potential selection bias induced by the textual information is ignored. In this work, we pro…

Cited by 0SourceScholar