← Search

Ru Wang

3 accepted papers

2026

SELF-HARMONY: LEARNING TO HARMONIZE SELF-SUPERVISION AND SELF-PLAY IN TEST-TIME REINFORCEMENT LEARNING

ICLR 2026poster

Test-time reinforcement learning (TTRL) offers a label-free paradigm for adapting models using only synthetic signals at inference, but its success hinges on constructing reliable learning signals. Standard approaches such as majority voting often collapse to spurious yet popular answers. We introdu…

Cited by 0SourceScholar
2026

Towards High-resolution and Disentangled Reference-based Sketch Colorization

CVPR 2026

Sketch colorization models have been widely studied to automate and assist in the creation of animation frames and digital illustrations. However, current methods are still not satisfactory for industrial standard applications in high-resolution synthesis and precise controllability of details. To f

Cited by 0SourcecodeScholar
2025

Semantic-Aware Prompt Learning for Multimodal Sarcasm Detection

ICASSP 2025accepted

Multimodal sarcasm detection aims to identify whether utterances express sarcastic intentions contrary to their literal meaning based on multimodal information. However, existing methods fail to explore the model’s "ability to understand" the semantics expressed by sentences in the image context fro…

Cited by 0SourceScholar