← Search

Bochen Wang

2 accepted papers

2025

AirCache: Activating Inter-modal Relevancy KV Cache Compression for Efficient Large Vision-Language Model Inference

ICCV 2025poster

Recent advancements in Large Visual Language Models (LVLMs) have gained significant attention due to their remarkable reasoning capabilities and proficiency in generalization. However, processing a large number of visual tokens and generating long-context outputs impose substantial computational ove…

Cited by 0SourcePDFScholar
2024

IVTP: Instruction-guided Visual Token Pruning for Large Vision-Language Models

ECCV 2024poster

"Inspired by the remarkable achievements of Large Language Models (LLMs), Large Vision-Language Models (LVLMs) have likewise experienced significant advancements. However, the increased computational cost and token budget occupancy associated with lengthy visual tokens pose significant challenge to…

Cited by 3SourcePDFScholar