← Search

Jingwen Zhao

2 accepted papers

2025

Hallucination Reduction in Video-Language Models via Hierarchical Multimodal Consistency

IJCAI 2025

The rapid advancement of large language models (LLMs) has led to the widespread adoption of video-language models (VLMs) across various domains. However, VLMs are often hindered by their limited semantic discrimination capability, exacerbated by the limited diversity and biased sample distribution o

Cited by 0SourcePDFScholar
2017

3D tracking swimming fish school with learned kinematic model using LSTM network

ICASSP 2017accepted

This paper proposes a reliable 3D fish tracking method using a novel master-slave camera setup. Instead of conventional dynamic models that rely on prior knowledge about target kinematics, the proposed method learns the kinematic model with a Long Short-Term Memory (LSTM) network. On this basis, the…

Cited by 0SourceScholar