← Search

Yixin Yang

14 accepted papers

2026

Augmented Radiance Field: A General Framework for Enhanced Gaussian Splatting

ICLR 2026poster

Due to the real-time rendering performance, 3D Gaussian Splatting (3DGS) has emerged as the leading method for radiance field reconstruction. However, its reliance on spherical harmonics for color encoding inherently limits its ability to separate diffuse and specular components, making it challengi…

Cited by 0SourceScholar
2026

RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection

AAAI 2026technical

Data selection for instruction tuning is crucial for improving the performance of large language models (LLMs) while reducing training costs. In this paper, we propose Refined Contribution Measurement with In-Context Learning (RICo), a novel gradient-free method that quantifies the fine-grained cont

Cited by 0SourcePDFScholar
2026

STCDiT: Spatio-Temporally Consistent Diffusion Transformer for High-Quality Video Super-Resolution

CVPR 2026

We present STCDiT, a video super-resolution framework built upon a pre-trained video diffusion model, aiming to restore structurally faithful and temporally stable videos from degraded inputs, even under complex camera motions. The main challenges lie in maintaining temporal stability during reconst

Cited by 0SourceScholar
2025

Beyond Single Frames: Can LMMs Comprehend Implicit Narratives in Comic Strip?

EMNLP 2025

Large Multimodal Models (LMMs) have demonstrated strong performance on vision-language benchmarks, yet current evaluations predominantly focus on single-image reasoning. In contrast, real-world scenarios always involve understanding sequences of images. A typical scenario is comic strips understandi

Cited by 0SourcePDFScholar
2025

Event-guided HDR Reconstruction with Diffusion Priors

ICCV 2025poster

Events provide High Dynamic Range (HDR) intensity change which can guide Low Dynamic Range (LDR) image for HDR reconstruction. However, events only provide temporal intensity differences and it is still ill-posed in over-/under-exposed areas due to missing initial reference brightness and color info…

2025

Polarimetric Neural Field via Unified Complex-Valued Wave Representation

ICCV 2025poster

Polarization has found applications in various computer vision tasks by providing additional physical cues. However, due to the limitations of current imaging systems, polarimetric parameters are typically stored in discrete form, which is non-differentiable and limits their applicability in polariz…

Cited by 0SourcePDFScholar
2024

Can Large Multimodal Models Uncover Deep Semantics Behind Images?

ACL 2024findings

Understanding the deep semantics of images is essential in the era dominated by social media. However, current research works primarily on the superficial description of images, revealing a notable deficiency in the systematic investigation of the inherent deep semantics. In this work, we introduce…

2024

ColorMNet: A Memory-based Deep Spatial-Temporal Feature Propagation Network for Video Colorization

ECCV 2024poster

"How to effectively explore spatial-temporal features is important for video colorization. Instead of stacking multiple frames along the temporal dimension or recurrently propagating estimated features that will accumulate errors or cannot explore information from far-apart frames, we develop a memo…

2024

Latency Correction for Event-guided Deblurring and Frame Interpolation

CVPR 2024poster

Event cameras with their high temporal resolution dynamic range and low power consumption are particularly good at time-sensitive applications like deblurring and frame interpolation. However their performance is hindered by latency variability especially under low-light conditions and with fast-mov…

Cited by 9SourcePDFScholar
2024

Sparse Bayesian Synthetic Aperture Processing Based DOA Estimation with Deformed Towed Arrays

ICASSP 2024accepted

In this paper, we present a new synthetic aperture method for direction-of-arrival (DOA) estimation using a passive towed sonar array that is deformed during platform maneuver. With certain prior knowledge of source-array geometry, we propose to find the optimal maximum likelihood estimates of DOAs…

Cited by 0SourceScholar
2023

Coherent Event Guided Low-Light Video Enhancement

ICCV 2023poster

With frame-based cameras, capturing fast-moving scenes without suffering from blur often comes at the cost of low SNR and low contrast. Worse still, the photometric constancy that enhancement techniques heavily relied on is fragile for frames with short exposure. Event cameras can record brightness…

Cited by 32PDFcodeScholar
2023

Learning Event Guided High Dynamic Range Video Reconstruction

CVPR 2023poster

Limited by the trade-off between frame rate and exposure time when capturing moving scenes with conventional cameras, frame based HDR video reconstruction suffers from scene-dependent exposure ratio balancing and ghosting artifacts. Event cameras provide an alternative visual representation with a m…

2021

EvIntSR-Net: Event Guided Multiple Latent Frames Reconstruction and Super-Resolution

ICCV 2021poster

An event camera detects the scene radiance changes and sends a sequence of asynchronous event streams with high dynamic range, high temporal resolution, and low latency. However, the spatial resolution of event cameras is limited as a trade-off for these outstanding properties. To reconstruct high-r…

Cited by 54PDFScholar