← Search

Bingjie WANG

4 accepted papers

2025

Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach

CVPR 2025poster

Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated remarkable progress in visual understanding. This impressive leap raises a compelling question: how can language models, initially trained solely on linguistic data, effectively interpret and process visual content? Th…

Cited by 4SourcePDFScholar
2024

On the Robustness of Neural-Enhanced Video Streaming against Adversarial Attacks

AAAI 2024technical

The explosive growth of video traffic on today's Internet promotes the rise of Neural-enhanced Video Streaming (NeVS), which effectively improves the rate-distortion trade-off by employing a cheap neural super-resolution model for quality enhancement on the receiver side. Missing by existing work, w…

Cited by 10SourcePDFScholar
2024

Towards Safe Concept Transfer of Multi-Modal Diffusion via Causal Representation Editing

NeurIPS 2024poster

Recent advancements in vision-language-to-image (VL2I) diffusion generation have made significant progress. While generating images from broad vision-language inputs holds promise, it also raises concerns about potential misuse, such as copying artistic styles without permission, which could have le…

Cited by 0SourcePDFScholar
2023

Towards Test-Time Refusals via Concept Negation

NeurIPS 2023poster

Generative models produce unbounded outputs, necessitating the use of refusal techniques to confine their output space. Employing generative refusals is crucial in upholding the ethical and copyright integrity of synthesized content, particularly when working with widely adopted diffusion models. "C…

Cited by 5SourcePDFScholar