← Search

Naotaka Kawata

2 accepted papers

2026

Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models

AAAI 2026technical

Contrastive pre-trained vision-language models, such as CLIP, demonstrate strong generalization abilities in zero-shot classification by leveraging embeddings extracted from image and text encoders. This paper aims to robustly fine-tune these vision-language models on in-distribution (ID) data witho

Cited by 0SourcePDFScholar
2024

Talking Face Generation for Impression Conversion Considering Speech Semantics

ICASSP 2024accepted

This study investigates the talking face generation method to convert a speaker’s video to give a target impression, such as “favorable” or “considerate”. Such an impression conversion method needs to consider the input speech semantics because they affect the impression of a speaker’s video along w…

Cited by 0SourceScholar