← Search

Fangming Feng

2 accepted papers

2026

Scene-Aware Spatiotemporal Generalization: Towards Robust Temporal Action Detection Across Domains

AAAI 2026technical

Temporal Action Detection (TAD) aims to identify specific actions in long, untrimmed videos by determining their start, end times and categories, yet existing models suffer from performance degradation under out-of-distribution scenarios due to unrealistic i.i.d. assumptions. While domain generaliza

Cited by 0SourcePDFScholar
2025

Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision

ICLR 2025poster

Prompt learning has demonstrated promising results in fine-tuning pre-trained multimodal models. However, the performance improvement is limited when applied to more complex and fine-grained tasks. The reason is that most existing methods directly optimize the parameters involved in the prompt gener…