← Search

Zhixin Zhang

8 accepted papers

2026

Group-aware Multiscale Ensemble Learning for Test-Time Multimodal Sentiment Analysis

AAAI 2026technical

Multi-modal Sentiment Analysis (MSA) enables machines to perceive human sentiments by integrating multiple modalities such as text, video, and audio. Despite recent progress, most existing methods assume distribution consistency between training and test data—a condition rarely met in real-world sce

Cited by 0SourcePDFScholar
2026

Monitoring LLM-based Multi-Agent Systems Against Corruptions via Node Evaluation

ICML 2026poster

Large Language Model (LLM)-based Multi-Agent Systems (MAS) have become a popular paradigm of AI applications. However, trustworthiness issues in MAS remain a critical concern. Unlike challenges in single-agent systems, MAS involve more complex communication processes, making them susceptible to corr…

Cited by 0SourceScholar
2026

UniAPO: Unified Multimodal Automated Prompt Optimization

AAAI 2026technical

Prompting is fundamental to unlocking the full potential of large language models. To automate and enhance this process, automatic prompt optimization (APO) has been developed, demonstrating effectiveness primarily in text-only input scenarios. However, extending existing APO methods to multimodal t

Cited by 0SourcePDFScholar
2025

Debt Collection Negotiations with Large Language Models: An Evaluation System and Optimizing Decision Making with Multi-Agent

ACL 2025finding

Debt collection negotiations (DCN) are vital for managing non-performing loans (NPLs) and reducing creditor losses. Traditional methods are labor-intensive, while large language models (LLMs) offer promising automation potential. However, prior systems lacked dynamic negotiation and real-time decisi…

2025

PL-VIWO: A Lightweight and Robust Point-Line Monocular Visual Inertial Wheel Odometry

IROS 2025

This paper presents a novel tightly coupled Filter-based monocular visual-inertial-wheel odometry (VIWO) system for ground robots, designed to deliver accurate and robust localization in long-term complex outdoor navigation scenarios. As an external sensor, the camera enhances localization performan

Cited by 2SourcecodeScholar
2025

Stackelberg Self-Annotation: A Robust Approach to Data-Efficient LLM Alignment

NeurIPS 2025poster

Aligning large language models (LLMs) with human preferences typically demands vast amounts of meticulously curated data, which is both expensive and prone to labeling noise. We propose Stackelberg Game Preference Optimization (SGPO), a robust alignment framework that models alignment as a two-playe…

Cited by 0SourceScholar
2024

Online Vectorized HD Map Construction using Geometry

ECCV 2024poster

"Online vectorized High-Definition (HD) map construction is critical for downstream prediction and planning. Recent efforts have built strong baselines for this task, however, geometric shapes and relations of instances in road systems are still under-explored, such as parallelism, perpendicular, re…

2021

TransForensics: Image Forgery Localization With Dense Self-Attention

ICCV 2021poster

Nowadays advanced image editing tools and technical skills produce tampered images more realistically, which can easily evade image forensic systems and make authenticity verification of images more difficult. To tackle this challenging problem, we introduce TransForensics, a novel image forgery loc…

Cited by 65PDFScholar