← Search

Jiaxing Zhao

2 accepted papers

2026

Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness

AAAI 2026technical

Facial expression captioning has found widespread application across various domains. Recently, the emergence of video Multimodal Large Language Models (MLLMs) has shown promise in general video understanding tasks. However, describing facial expressions within videos poses two major challenges for

Cited by 0SourcePDFScholar
2026

Weaving in the Clouds: Achieving Synergistic Collaboration among LLM Agents via Federated Learning

ICML 2026poster

Multi-Agent Systems (MAS) powered by Large Language Models (LLMs) have recently become a strong paradigm for solving complex workflow-structured tasks through expert collaboration. However, the data that make such collaboration effective are typically distributed across organizations and cannot be c…

Cited by 0SourceScholar