← Search

Jiapeng Shi

1 accepted papers

2026

VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding

ICML 2026poster

Recent advancements in Video Large Language Models (Video LLMs) have demonstrated impressive results, yet existing approaches handle either temporal or spatial dimension in isolation, struggling in the analysis of complex events that require spatial-temporal integration. To bridge this gap, we propo…

Cited by 5SourceScholar