← Search

Taiga Yamane

3 accepted papers

2026

Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models

AAAI 2026technical

Contrastive pre-trained vision-language models, such as CLIP, demonstrate strong generalization abilities in zero-shot classification by leveraging embeddings extracted from image and text encoders. This paper aims to robustly fine-tune these vision-language models on in-distribution (ID) data witho

Cited by 0SourcePDFScholar
2025

MVTrajecter: Multi-View Pedestrian Tracking with Trajectory Motion Cost and Trajectory Appearance Cost

ICCV 2025poster

Multi-View Pedestrian Tracking (MVPT) aims to track pedestrians in the form of a bird's eye view occupancy map from multi-view videos. End-to-end methods that detect and associate pedestrians within one model have shown great progress in MVPT. The motion and appearance information of pedestrians is…

Cited by 4SourcePDFScholar
2023

Leveraging Language Embeddings for Cross-Lingual Self-Supervised Speech Representation Learning

ICASSP 2023accepted

In this paper, we propose novel cross-lingual self-supervised speech representation learning methods that explicitly consider language information. Cross-lingual self-supervised speech representation learning has been studied to make effective use of diverse data in various languages. Previous metho…

Cited by 0SourceScholar