← Search

Anlong Ming

23 accepted papers

2026

Regression over Classification: Assessing Image Aesthetics via Multimodal Large Language Models

AAAI 2026technical

Image Aesthetics Assessment (IAA) evaluates visual quality through user-centered perceptual analysis and can guide various applications. Recent advances in Multimodal Large Language Models (MLLMs) have sparked interest in adapting them for IAA. However, two critical limitations persist in applying M

Cited by 0SourcePDFScholar
2026

Thinking Aesthetics Assessment of Image Color Temperature: Models, Datasets and Benchmarks

AAAI 2026technical

Color temperature, as a crucial attribute influencing image color, plays a critical role in Image Aesthetics Assessment (IAA). Yet, within the existing IAA field, little light has been shed on assessing the aesthetic quality of image color temperature. To bridge this gap, we introduce a new task: Im

Cited by 0SourcePDFScholar
2025

Rethinking Personalized Aesthetics Assessment: Employing Physique Aesthetics Assessment as An Exemplification

CVPR 2025highlight

The Personalized Aesthetics Assessment (PAA) aims to accurately predict an individual's unique perception of aesthetics. With the surging demand for customization, PAA enables applications to generate personalized outcomes by aligning with individual aesthetic preferences. The prevailing PAA paradig…

2024

AK4Prompts: Aesthetics-driven Automatically Keywords-Ranking for Prompts in Text-To-Image Models

IJCAI 2024poster

Current text-to-image synthesis (TIS) models have demonstrated the ability to generate high-fidelity images based on textual prompts. However, the efficacy of these models heavily relies on the keywords present in the prompts, and there is a dearth of objective analysis regarding how different keywo…

2024

ELTA: An Enhancer against Long-Tail for Aesthetics-oriented Models

ICML 2024poster

Real-world datasets often exhibit long-tailed distributions, compromising the generalization and fairness of learning-based models. This issue is particularly pronounced in Image Aesthetics Assessment (IAA) tasks, where such imbalance is difficult to mitigate due to a severe distribution mismatch be…

Cited by 4SourcePDFScholar
2024

M2Beats: When Motion Meets Beats in Short-form Videos

IJCAI 2024poster

In recent years, short-form videos have gained popularity and the editing of these videos, particularly when motion is synchronized with music, is highly favored due to its beat-matching effect. However, detecting motion rhythm poses a significant challenge as it is influenced by multiple factors…

2024

Rethinking No-reference Image Exposure Assessment from Holism to Pixel: Models, Datasets and Benchmarks

NeurIPS 2024poster

The past decade has witnessed an increasing demand for enhancing image quality through exposure, and as a crucial prerequisite in this endeavor, Image Exposure Assessment (IEA) is now being accorded serious attention. However, IEA encounters two persistent challenges that remain unresolved over the…

2023

ICDA: Illumination-Coupled Domain Adaptation Framework for Unsupervised Nighttime Semantic Segmentation

IJCAI 2023poster

The performance of nighttime semantic segmentation has been significantly improved thanks to recent unsupervised methods. However, these methods still suffer from complex domain gaps, i.e., the challenging illumination gap and the inherent dataset gap. In this paper, we propose the illumination-coup…

2023

Thinking Image Color Aesthetics Assessment: Models, Datasets and Benchmarks

ICCV 2023poster

We present a comprehensive study on a new task named image color aesthetics assessment (ICAA), which aims to assess color aesthetics based on human perception. ICAA is important for various applications such as imaging measurement and image analysis. However, due to the highly diverse aesthetic pref…

Cited by 22PDFcodeScholar
2023

Unknown Sniffer for Object Detection: Don't Turn a Blind Eye to Unknown Objects

CVPR 2023poster

The recently proposed open-world object and open-set detection have achieved a breakthrough in finding never-seen-before objects and distinguishing them from known ones. However, their studies on knowledge transfer from known classes to unknown ones are not deep enough, resulting in the scanty capab…

2023

WBFlow: Few-shot White Balance for sRGB Images via Reversible Neural Flows

IJCAI 2023poster

The sRGB white balance methods aim to correct the nonlinear color cast of sRGB images without accessing raw values. Although existing methods have achieved increasingly better results, their generalization to sRGB images from multiple cameras is still under explored. In this paper, we propose…

2022

Fast Road Segmentation via Uncertainty-aware Symmetric Network

ICRA 2022poster

The high performance of RGB-D based road segmentation methods contrasts with their rare application in commercial autonomous driving, which is owing to two reasons: 1) the prior methods cannot achieve high inference speed and high accuracy in both ways; 2) the different properties of RGB and depth d…

Cited by 48SourcecodeScholar
2022

Monocular Depth Distribution Alignment with Low Computation

ICRA 2022poster

The performance of monocular depth estimation generally depends on the amount of parameters and computational cost. It leads to a large accuracy contrast between light-weight networks and heavy-weight networks, which limits their application in the real world. In this paper, we model the majority of…

Cited by 14SourcecodeScholar
2022

Rethinking Image Aesthetics Assessment: Models, Datasets and Benchmarks

IJCAI 2022poster

Challenges in image aesthetics assessment (IAA) arise from that images of different themes correspond to different evaluation criteria, and learning aesthetics directly from images while ignoring the impact of theme variations on human visual perception inhibits the further development of IAA; howev…

2022

Transfer Learning for Color Constancy via Statistic Perspective

AAAI 2022technical

Color Constancy aims to correct image color casts caused by scene illumination. Recently, although the deep learning approaches have remarkably improved on single-camera data, these models still suffer from the seriously insufficient data problem, resulting in shallow model capacity and degradation…

Cited by 14SourcePDFScholar
2021

MT-ORL: Multi-Task Occlusion Relationship Learning

ICCV 2021poster

Retrieving occlusion relation among objects in a single image is challenging due to sparsity of boundaries in image. We observe two key issues in existing works: firstly, lack of an architecture which can exploit the limited amount of coupling in the decoder stage between the two subtasks, namely oc…

Cited by 8PDFcodeScholar
2019

A Novel Super-resolution Method Based on Patch Reconstruction with Simk Clustering and Nonlinear Mapping

ICASSP 2019accepted

In this paper, we propose a patch-wise super-resolution (SR) method that combines an external-sample classification tree and a nonlinear-mapping learning stage to simultaneously guarantees reconstruction quality and speed at the stage of patch representation and mapping. We use the low-resolution (L…

Cited by 0SourceScholar
2019

Occlusion-Shared and Feature-Separated Network for Occlusion Relationship Reasoning

ICCV 2019poster

Occlusion relationship reasoning demands closed contour to express the object, and orientation of each contour pixel to describe the order relationship between objects. Current CNN-based methods neglect two critical issues of the task: (1) simultaneous existence of the relevance and distinction for…

Cited by 33PDFcodeScholar
2019

Spatio-spectral Modulation Using a Binary Photomask for Compressive Chromotomography

ICASSP 2019accepted

Recent advances in compressive spectral imagers have demonstrated the potential of spatio-spectral modulation (SSM) for improved reconstruction performance. Existing SSM techniques, however, use either a color filter array or a complex optical arrangement, both of which can only provide limited modu…

Cited by 0SourceScholar