← Search

Mingcan Ma

4 accepted papers

2025

FICGen: Frequency-Inspired Contextual Disentanglement for Layout-driven Degraded Image Generation

ICCV 2025poster

Layout-to-image (L2I) generation has exhibited promising results in natural domains, but suffers from limited generative fidelity and weak alignment with user-provided layouts when applied to degraded scenes (i.e., low-light, underwater). We primarily attribute these limitations to the "contextual i…

Cited by 0SourcePDFScholar
2025

FreeGen: Bridging Visual-Linguistic Discrepancies Towards Diffusion-based Pixel-level Data Synthesis

AAAI 2025technical

Text-to-image diffusion model has inspired research into text-to-data synthesis without human intervention, where spatial attentions correlated with semantic entities in text prompts are primarily interpreted as pseudo-masks. However, these vannila attentions often deliver visual-linguistic discrepa…

2022

Pyramid Grafting Network for One-Stage High Resolution Saliency Detection

CVPR 2022poster

Recent salient object detection (SOD) methods based on deep neural network have achieved remarkable performance. However, most of existing SOD models designed for low-resolution input perform poorly on high-resolution images due to the contradiction between the sampling depth and the receptive field…

Cited by 134PDFcodeScholar