← Search

Jiawang Liu

2 accepted papers

2025

SiQA: A Large Multi-Modal Question Answering Model for Structured Images Based on RAG

ICASSP 2025accepted

Existing Large Multimodal Models (LMMs) demonstrate excellent performance in handling visual tasks in everyday scenarios. However, they still face challenges in understanding structured images, such as flowcharts and organizational charts, which are characterized by text-rich and complex hierarchica…

Cited by 0SourceScholar
2023

Attention Based Relation Network for Facial Action Units Recognition

ICASSP 2023accepted

Facial action unit (AU) recognition is essential to facial expression analysis. Since there are highly positive or negative correlations between AUs, some existing AU recognition works have focused on modeling AU relations. However, previous relationship-based approaches typically embed predefined r…

Cited by 0SourceScholar