← Search

Chengbin Quan

4 accepted papers

2024

LLaMA-Excitor: General Instruction Tuning via Indirect Feature Interaction

CVPR 2024poster

Existing methods to fine-tune LLMs like Adapter Prefix-tuning and LoRA which introduce extra modules or additional input sequences to inject new skills or knowledge may compromise the innate abilities of LLMs. In this paper we propose LLaMA-Excitor a lightweight method that stimulates the LLMs' pote…

Cited by 3SourcePDFScholar
2024

Language-aware Visual Semantic Distillation for Video Question Answering

CVPR 2024poster

Significant advancements in video question answering (VideoQA) have been made thanks to thriving large image-language pretraining frameworks. Although these image-language models can efficiently represent both video and language branches they typically employ a goal-free vision perception process an…

Cited by 3SourcePDFScholar
2024

Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment based on Multi-Scale Aggregation and Anthropic Prior Knowledge

CVPR 2024poster

Teeth localization segmentation and labeling in 2D images have great potential in modern dentistry to enhance dental diagnostics treatment planning and population-based studies on oral health. However general instance segmentation frameworks are incompetent due to 1) the subtle differences between s…

Cited by 3SourcePDFScholar
2022

Delving into Sequential Patches for Deepfake Detection

NeurIPS 2022accept

Recent advances in face forgery techniques produce nearly visually untraceable deepfake videos, which could be leveraged with malicious intentions. As a result, researchers have been devoted to deepfake detection. Previous studies have identified the importance of local low-level cues and temporal i…

Cited by 66SourcePDFScholar