← Search

Ziyang Xie

4 accepted papers

2025

Vid2Sim: Realistic and Interactive Simulation from Video for Urban Navigation

CVPR 2025poster

Sim-to-real gap has long posed a significant challenge for robot learning in simulation, preventing the deployment of learned models in the real world. Previous work has primarily focused on domain randomization and system identification to mitigate this gap. However, these methods are often limited…

Cited by 4SourcePDFScholar
2024

Frozen Transformers in Language Models Are Effective Visual Encoder Layers

ICLR 2024spotlight

This paper reveals that large language models (LLMs), despite being trained solely on text data, are surprisingly}strong encoders for purely visual tasks in the absence of language. Even more intriguingly, this can be achieved by a simple yet previously overlooked strategy -- employing a frozen tran…