← Search

Yongjian Deng

22 accepted papers

2026

AIR-DR: Adaptive Image Retargeting with Instance Relocation and Dual-guidance Repainting

AAAI 2026technical

Image retargeting aims to adjust the aspect ratio of images to accommodate various display devices. While existing methods consider both foreground semantics and background inpainting, their Seam-carving-based framework is inherently destructive, often compromising the structural integrity of foregr

Cited by 0SourcePDFScholar
2026

DF^2-VB: Dual-level Fuzzy Fusion with View-specific Boosting for Multi-view Multi-label Classification

CVPR 2026

Multi-view multi-label classification (MVMLC) aims to utilize both consensus and complementarity information to predict potentially relevant labels for samples. Existing MVMLC approaches typically focus on either feature-level fusion, which integrates complementary features for more expressive repre

Cited by 0SourceScholar
2026

Event-Based Motion Deblurring Using Task-Oriented 3D Gaussian Event Representations

CVPR 2026

Event-based motion deblurring has attracted increasing attention, as the high temporal resolution of event cameras provides motion cues unavailable to conventional RGB sensors, thereby enabling more effective deblurring. In real-world scenes, motion blur is often complex and nonlinear, with differen

Cited by 0SourceScholar
2026

One-Shot Flow, Any-Time Frame: A Bidirectional Warping Framework for Event-Based Video Frame Interpolation

CVPR 2026

Video Frame Interpolation (VFI) is a crucial task in video processing. Flow-based methods, despite their success, are constrained by a fundamental dilemma: forward warping is efficient but prone to artifacts, while backward warping yields higher quality at a significant computational cost, especiall

Cited by 0SourcecodeScholar
2025

CFDM: Contrastive Fusion and Disambiguation for Multi-View Partial-Label Learning

AAAI 2025technical

When dealing with multi-view data, the heterogeneity of data attributes across different views often leads to label ambiguity. To effectively address this challenge, this paper designs a Multi-View Partial-Label Learning (MVPLL) framework, where each training instance is described by multiple view f…

Cited by 0SourcePDFScholar
2025

EPA: Boosting Event-based Video Frame Interpolation with Perceptually Aligned Learning

NeurIPS 2025poster

Event cameras, with their capacity to provide high temporal resolution information between frames, are increasingly utilized for video frame interpolation (VFI) in challenging scenarios characterized by high-speed motion and significant occlusion. However, prevalent issues of blur and distortion wit…

Cited by 0SourceScholar
2025

ESEG: Event-Based Segmentation Boosted by Explicit Edge-Semantic Guidance

AAAI 2025technical

Event-based semantic segmentation (ESS) has attracted researchers' attention recently, as event cameras can solve problems such as under/over-exposure or motion blur that are difficult for RGB cameras to handle. However, event data are noisy and sparse, resulting in difficulties for the model to loc…

2025

Enhance Multi-View Classification Through Multi-Scale Alignment and Expanded Boundary

ICLR 2025poster

Multi-view classification aims at unifying the data from multiple views to complementarily enhance the classification performance. Unfortunately, two major problems in multi-view data are damaging model performance. The first is feature heterogeneity, which makes it hard to fuse features from differ…

Cited by 0SourcePDFScholar
2025

Graph Consistency and Diversity Measurement for Federated Multi-View Clustering

AAAI 2025technical

Federated Multi-View Clustering (FMVC) aims to learn a global clustering model from heterogeneous data distributed across different devices, where each device only stores one view of all clustering samples. The key to deal with such problem lies in how to effectively fuse these heterogeneous samples…

Cited by 0SourcePDFScholar
2025

Know Where You Are From: Event-Based Segmentation via Spatio-Temporal Propagation

AAAI 2025technical

Event cameras have gained attention in segmentation due to their higher temporal resolution and dynamic range compared to traditional cameras. However, they struggle with issues like lack of color perception and triggering only at motion edges, making it hard to distinguish objects with similar cont…

2025

MSV-PCT: Multi-Sparse-View Enhanced Transformer Framework for Salient Object Detection in Point Clouds

AAAI 2025technical

Salient object detection (SOD) methods for 2D images have great significance in the field of human-computer interaction (HCI). However, as a common data format in HCI, the SOD research in the form of 3D point cloud data remains limited. Previous works commonly treat this task as point cloud segmenta…

Cited by 0SourcePDFScholar
2025

Multi-View Multi-Label Classification via View-Label Matching Selection

AAAI 2025technical

In multi-view multi-label classification (MVML), each object is described by several heterogeneous views while annotated with multiple related labels. The key to learn from such complicate data lies in how to fuse cross-view features and explore multi-label correlations, while accordingly obtain cor…

Cited by 0SourcePDFScholar
2024

A Dynamic GCN with Cross-Representation Distillation for Event-Based Learning

AAAI 2024technical

Recent advances in event-based research prioritize sparsity and temporal precision. Approaches learning sparse point-based representations through graph CNNs (GCN) become more popular. Yet, these graph techniques hold lower performance than their frame-based counterpart due to two issues: (i) Biased…

Cited by 8SourcePDFScholar
2024

A Motion-aware Spatio-temporal Graph for Video Salient Object Ranking

NeurIPS 2024poster

Video salient object ranking aims to simulate the human attention mechanism by dynamically prioritizing the visual attraction of objects in a scene over time. Despite its numerous practical applications, this area remains underexplored. In this work, we propose a graph model for video salient object…

2024

Prune and Repaint: Content-Aware Image Retargeting for any Ratio

NeurIPS 2024poster

Image retargeting is the task of adjusting the aspect ratio of images to suit different display devices or presentation environments. However, existing retargeting methods often struggle to balance the preservation of key semantics and image quality, resulting in either deformation or loss of import…

2024

SAM-Event-Adapter: Adapting Segment Anything Model for Event-RGB Semantic Segmentation

ICRA 2024poster

Semantic segmentation, a fundamental visual task ubiquitously employed in sectors ranging from transportation and robotics to healthcare, has always captivated the research community. In the wake of rapid advancements in large model research, the foundation model for semantic segmentation tasks, ter…

Cited by 11SourceScholar
2024

Video Frame Interpolation via Direct Synthesis with the Event-based Reference

CVPR 2024poster

Video Frame Interpolation (VFI) has witnessed a surge in popularity due to its abundant downstream applications. Event-based VFI (E-VFI) has recently propelled the advancement of VFI. Thanks to the high temporal resolution benefits event cameras can bridge the informational void present between succ…

Cited by 6SourcePDFScholar
2023

TokenHPE: Learning Orientation Tokens for Efficient Head Pose Estimation via Transformers

CVPR 2023poster

Head pose estimation (HPE) has been widely used in the fields of human machine interaction, self-driving, and attention estimation. However, existing methods cannot deal with extreme head pose randomness and serious occlusions. To address these challenges, we identify three cues from head images, na…

2022

VMV-GCN: Volumetric Multi-View Based Graph CNN for Event Stream Classification

RA-L 2022

Event cameras can perceive pixel-level brightness changes to output asynchronous event streams, and have notable advantages in high temporal resolution, high dynamic range and low power consumption for challenging vision tasks. To apply existing learning models on event data, many researchers integr

Cited by 54SourceScholar