← Search

Longwen Gao

6 accepted papers

2025

Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models

EMNLP 2025

Object hallucinations in Large Vision-Language Models (LVLMs) significantly impede their real-world applicability. As the primary component for accurately interpreting visual information, the choice of visual encoder is pivotal. We hypothesize that the diverse training paradigms employed by differen

2025

MX-Font++: Mixture of Heterogeneous Aggregation Experts for Few-shot Font Generation

ICASSP 2025accepted

Few-shot Font Generation (FFG) aims to create new font libraries using limited reference glyphs, with crucial applications in digital accessibility and equity for low-resource languages, especially in multilingual artificial intelligence systems. Although existing methods have shown promising perfor…

Cited by 0SourceScholar
2023

Mining and Applying Composition Knowledge of Dance Moves for Style-Concentrated Dance Generation

AAAI 2023technical

Choreography refers to creation of dance motions according to both music and dance knowledge, where the created dances should be style-specific and consistent. However, most of the existing methods generate dances using the given music as the only reference, lacking the stylized dancing knowledge, n…

2023

Video Compression Artifact Reduction by Fusing Motion Compensation and Global Context in a Swin-CNN Based Parallel Architecture

AAAI 2023technical

Video Compression Artifact Reduction aims to reduce the artifacts caused by video compression algorithms and improve the quality of compressed video frames. The critical challenge in this task is to make use of the redundant high-quality information in compressed frames for compensation as much as p…

2021

GIF Thumbnails: Attract More Clicks to Your Videos

AAAI 2021technical

With the rapid increase of mobile devices and online media, more and more people prefer posting/viewing videos online. Generally, these videos are presented on video streaming sites with image thumbnails and text titles. While facing huge amounts of videos, a viewer clicks through a certain video wi…