← Search

Xiaomeng Wang

5 accepted papers

2025

StarGen: A Spatiotemporal Autoregression Framework with Video Diffusion Model for Scalable and Controllable Scene Generation

CVPR 2025poster

Recent advances in large reconstruction and generative models have significantly improved scene reconstruction and novel view generation. However, due to compute limitations, each inference with these large models is confined to a small area, making long-range consistent scene generation challenging…

Cited by 1SourcePDFScholar
2025

Towards Rationality in Language and Multimodal Agents: A Survey

NAACL 2025long

This work discusses how to build more rational language and multimodal agents and what criteria define rationality in intelligent systems.Rationality is the quality of being guided by reason, characterized by decision-making that aligns with evidence and logical principles. It plays a crucial role i…

2024

A Peek into Token Bias: Large Language Models Are Not Yet Genuine Reasoners

EMNLP 2024main

This study introduces a hypothesis-testing framework to assess whether large language models (LLMs) possess genuine reasoning abilities or primarily depend on token bias. We go beyond evaluating LLMs on accuracy; rather, we aim to investigate their token bias in solving logical reasoning tasks. Spec…

2015

TRIC-track: Tracking by Regression With Incrementally Learned Cascades

ICCV 2015poster

This paper proposes a novel approach to part-based tracking by replacing local matching of an appearance model by direct prediction of the displacement between local image patches and part locations. We propose to use cascaded regression with incremental learning to track generic objects without any…

Cited by 35PDFScholar