← Search

Wenlong Chen

13 accepted papers

2026

Collaborative Enhancement of Large and Small Models for Question Answering via Dual Knowledge Transfer

AAAI 2026technical

Our statistical analysis reveals a complementary phenomenon between large language model-based question answering (QA) and small model-based QA. To facilitate dual knowledge transfer between these two paradigms, this paper introduces a collaborative enhancement method of large and small models for q

Cited by 0SourcePDFScholar
2025

Compact Memory for Continual Logistic Regression

NeurIPS 2025poster

Despite recent progress, continual learning still does not match the performance of batch training. To avoid catastrophic forgetting, we need to build compact memory of essential past knowledge, but no clear solution has yet emerged, even for shallow neural networks with just one or two layers. In t…

Cited by 0SourceScholar
2025

MobiExo: GPS-SLAM Fusion for Seamless Indoor-Outdoor Mobile Manipulation with Hand-Foot Coordination

IROS 2025

Teleoperation systems for mobile robots face significant challenges in achieving seamless coordination across dynamic environments. We present MobiExo, a teleoperation system that unlocks seamless indoor-outdoor mobile manipulation. Our approach tackles two fundamental challenges: robust cross-envir

Cited by 0SourcecodeScholar
2025

Prototype-based Optimal Transport for Out-of-Distribution Detection

IJCAI 2025

Detecting Out-of-Distribution (OOD) inputs is crucial for improving the reliability of deep neural networks in the real-world deployment. In this paper, inspired by the inherent distribution shift between in-distribution (ID) and OOD data, we propose a novel method that leverages optimal transport t

2025

Recurrent Memory for Online Interdomain Gaussian Processes

NeurIPS 2025poster

We propose a novel online Gaussian process (GP) model that is capable of capturing long-term memory in sequential data in an online learning setting. Our model, Online HiPPO Sparse Variational Gaussian Process (OHSVGP), leverages the HiPPO (High-order Polynomial Projection Operators) framework, whic…

Cited by 0SourceScholar
2025

Variational Uncertainty Decomposition for In-Context Learning

NeurIPS 2025poster

As large language models (LLMs) gain popularity in conducting prediction tasks in-context, understanding the sources of uncertainty in in-context learning becomes essential to ensuring reliability. The recent hypothesis of in-context learning performing predictive Bayesian inference opens the avenue…

Cited by 0SourceScholar
2024

Out-of-Distribution Detection for Learning-Based Chest X-Ray Diagnosis

ICASSP 2024accepted

Deep learning has shown prominence in chest radiography interpretation, which is critical in evaluating various lung and chest diseases, such as pneumonia, emphysema, and tuberculosis. Deploying machine learning model, it is important to detect out-of-distribution (OOD) inputs, which are distinct fr…

Cited by 0SourceScholar
2024

Parallel Ranking of Ads and Creatives in Real-Time Advertising Systems

AAAI 2024technical

Creativity is the heart and soul of advertising services. Effective creatives can create a win-win scenario: advertisers each target users and achieve marketing objectives more effectively, users more quickly find products of interest, and platforms generate more advertising revenue. With the advent…

Cited by 2SourcePDFScholar
2023

Blending Advertising with Organic Content in E-commerce via Virtual Bids

AAAI 2023technical

It has become increasingly common that sponsored content (i.e., paid ads) and non-sponsored content are jointly displayed to users, especially on e-commerce platforms. Thus, both of these contents may interact together to influence their engagement behaviors. In general, sponsored content helps bran…

Cited by 7SourcePDFScholar
2021

A Gradient Based Strategy for Hamiltonian Monte Carlo Hyperparameter Optimization

ICML 2021spotlight

Hamiltonian Monte Carlo (HMC) is one of the most successful sampling methods in machine learning. However, its performance is significantly affected by the choice of hyperparameter values. Existing approaches for optimizing the HMC hyperparameters either optimize a proxy for mixing speed or consider…

Cited by 23SourcePDFScholar