← Search

Yuqi Chen

7 accepted papers

2026

GeniNav: Generative Model Driven Image-Goal Navigation via Imagination-Guided Consistency Flow Matching

CVPR 2026

Image-goal navigation driven by generative models has recently shown strong potential owing to their ability to perform multi-modal reasoning and stable learning in continuous control spaces. Despite their promise, current methods still face several fundamental limitations. Many rely on pre-built pr

Cited by 0SourcecodeScholar
2025

VMDT: Decoding the Trustworthiness of Video Foundation Models

NeurIPS 2025poster

As foundation models become more sophisticated, ensuring their trustworthiness becomes increasingly critical; yet, unlike text and image, the video modality still lacks comprehensive trustworthiness benchmarks. We introduce VMDT (Video-Modal DecodingTrust), the first unified platform for evaluating…

Cited by 0SourcecodeScholar
2024

Modeling Route Representation With Mixed-Scale Hierarchical Transformer

ICASSP 2024accepted

Modeling route representation aims to obtain contextual representations of an entire route for various traffic-related tasks. In reality, spatial-temporal data often exhibits multi-scale characteristics, which are utilized by many studies to enhance their performance. However, there is still a lack…

Cited by 0SourceScholar
2024

Surveying the Dead Minds: Historical-Psychological Text Analysis with Contextualized Construct Representation (CCR) for Classical Chinese

EMNLP 2024main

In this work, we develop a pipeline for historical-psychological text analysis in classical Chinese. Humans have produced texts in various languages for thousands of years; however, most of the computational literature is focused on contemporary languages and corpora. The emerging field of historica…

2023

Classifying Pathological Images Based on Multi-Instance Learning and End-to-End Attention Pooling

ICASSP 2023accepted

In order to address the issue that previous deep learning methods for classifying pathological images cannot adaptively learn features, we propose an end-to-end attention pooling method based on a multi-instance learning patch scoring model. Our method integrates feature extraction and classificatio…

Cited by 0SourceScholar
2023

ContiFormer: Continuous-Time Transformer for Irregular Time Series Modeling

NeurIPS 2023poster

Modeling continuous-time dynamics on irregular time series is critical to account for data evolution and correlations that occur continuously. Traditional methods including recurrent neural networks or Transformer models leverage inductive bias via powerful neural architectures to capture complex pa…

2022

A Span-level Bidirectional Network for Aspect Sentiment Triplet Extraction

EMNLP 2022main

Aspect Sentiment Triplet Extraction (ASTE) is a new fine-grained sentiment analysis task that aims to extract triplets of aspect terms, sentiments, and opinion terms from review sentences. Recently, span-level models achieve gratifying results on ASTE task by taking advantage of the predictions of a…