← Search

Wenshuo Peng

5 accepted papers

2026

PointSFDA: Source-Free Domain Adaptation for Point Cloud Completion

ICRA 2026poster

Point cloud completion is critical for autonomous driving and robotic perception, yet deep learning models often experience severe performance degradation under the domain gap between synthetic training and real-world data. While unsupervised domain adaptation (UDA) has been explored to mitigate thi…

2026

PyVision-RL: Forging Open Agentic Vision Models via RL

ICML 2026poster

Reinforcement learning for agentic multimodal models often suffers from interaction collapse, where models learn to reduce tool usage and multi-turn reasoning, limiting the benefits of agentic behavior. We introduce PyVision-RL, a reinforcement learning framework for open-weight multimodal models th…

Cited by 0SourceScholar
2026

SVBench: Evaluation of Video Generation Models on Social Reasoning

CVPR 2026

Recent text-to-video generation models have made remarkable progress in visual realism, motion fidelity, and text-video alignment, yet they still struggle to produce socially coherent behavior. Unlike humans, who readily infer intentions, beliefs, emotions, and social norms from brief visual cues, c

Cited by 0SourcecodeScholar
2024

Data Adaptive Traceback for Vision-Language Foundation Models in Image Classification

AAAI 2024technical

Vision-language foundation models have been incredibly successful in a wide range of downstream computer vision tasks using adaptation methods. However, due to the high cost of obtaining pre-training datasets, pairs with weak image-text correlation in the data exist in large numbers. We call them we…

Cited by 1SourcePDFScholar