← Search

Xinyang Liu

20 accepted papers

2026

DecFus: Decentralized Layer-wise Fusion with Dynamic Exploration and Exploitation

ICML 2026poster

Decentralized Federated Learning (DFL) enables collaborative model training across connected clients without a central server, effectively mitigating communication bottlenecks and avoiding the single point of failure in Centralized Federated Learning (CFL). However, existing DFL methods mostly focus…

Cited by 0SourceScholar
2026

SwarmNav: Swarm Robotics Navigation in Dynamic and Dense Environments Via Reinforcement Learning

ICRA 2026poster

Collision avoidance and navigation in dynamic and dense environments remain highly challenging for swarm robotics. To address this, we propose SwarmNav, a novel goal-region amplification navigation policy that leverages LiDAR-based position data to generate velocity commands guiding robots toward th…

Cited by 0Scholar
2026

Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct

ICLR 2026poster

Fast and high-quality language generation is the holy grail that people pursue in the age of AI. In this work, we introduce **Di**screte **Di**ffusion Divergence **Instruct** (**DiDi-Instruct**), a training-based method that initializes from a pre-trained diffusion large language model (dLLM) and di…

Cited by 0SourcecodeScholar
2025

Beyond Matryoshka: Revisiting Sparse Coding for Adaptive Representation

ICML 2025oral

Many large-scale systems rely on high-quality deep representations (embeddings) to facilitate tasks like retrieval, search, and generative modeling. Matryoshka Representation Learning (MRL) recently emerged as a solution for adaptive embedding lengths, but it requires full model retraining and suffe…

2025

Binarized Neural Network for Multi-spectral Image Fusion

CVPR 2025poster

Pan-sharpening technology refers to generating a high-resolution (HR) multi-spectral (MS) image with broad applications by fusing a low-resolution (LR) MS image and HR panchromatic (PAN) image. While deep learning approaches have shown impressive performance in pan-sharpening, they generally require…

Cited by 0SourcePDFScholar
2025

FLARE: Fast Large-Scale Autonomous Exploration Guided by Unknown Regions

RA-L 2025

Autonomous exploration is a critical foundation for unmanned aerial vehicle (UAV) applications such as search and rescue. However, existing methods typically focus only on known spaces or frontiers without considering unknown regions or providing further guidance for the global path, which results i

Cited by 2SourceScholar
2025

Physics-informed Neural Operator for Pansharpening

NeurIPS 2025poster

Over the past decades, pansharpening has contributed greatly to numerous remote sensing applications, with methods evolving from theoretically grounded models to deep learning approaches and their hybrids. Though promising, existing methods rarely address pansharpening through the lens of underlying…

Cited by 0SourceScholar
2025

SemiDFL: A Semi-Supervised Paradigm for Decentralized Federated Learning

AAAI 2025technical

Decentralized federated learning (DFL) realizes cooperative model training among connected clients without relying on a central server, thereby mitigating communication bottlenecks and eliminating the single-point failure issue present in centralized federated learning (CFL). Most existing work on…

2024

Byzantine-robust Decentralized Federated Learning via Dual-domain Clustering and Trust Bootstrapping

CVPR 2024poster

Decentralized federated learning (DFL) facilitates collaborative model training across multiple connected clients without a central coordination server thereby avoiding the single point of failure in traditional centralized federated learning (CFL). However DFL exhibits heightened susceptibility to…

Cited by 7SourcePDFScholar
2024

ETO:Efficient Transformer-based Local Feature Matching by Organizing Multiple Homography Hypotheses

NeurIPS 2024poster

We tackle the efficiency problem of learning local feature matching.Recent advancements have given rise to purely CNN-based and transformer-based approaches, each augmented with deep learning techniques. While CNN-based methods often excel in matching speed, transformer-based methods tend to provide…

Cited by 3SourcePDFScholar
2024

Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models

UAI 2024poster

For downstream applications of vision-language pre-trained models, there has been significant interest in constructing effective prompts. Existing works on prompt engineering, which either require laborious manual designs or optimize the prompt tuning as a point estimation problem, may fail to descr…

Cited by 3SourcePDFScholar
2023

Bayesian Progressive Deep Topic Model with Knowledge Informed Textual Data Coarsening Process

ICML 2023poster

Deep topic models have shown an impressive ability to extract multi-layer document latent representations and discover hierarchical semantically meaningful topics.However, most deep topic models are limited to the single-step generative process, despite the fact that the progressive generative proce…

Cited by 6SourcePDFScholar
2023

Context-guided Embedding Adaptation for Effective Topic Modeling in Low-Resource Regimes

NeurIPS 2023poster

Embedding-based neural topic models have turned out to be a superior option for low-resourced topic modeling. However, current approaches consider static word embeddings learnt from source tasks as general knowledge that can be transferred directly to the target task, discounting the dynamically cha…

2023

Multi-Modal Neural Radiance Field for Monocular Dense SLAM with a Light-Weight ToF Sensor

ICCV 2023poster

Light-weight time-of-flight (ToF) depth sensors are compact and cost-efficient, and thus widely used on mobile devices for tasks such as autofocus and obstacle detection. However, due to the sparse and noisy depth measurements, these sensors have rarely been considered for dense geometry reconstruct…

Cited by 32PDFcodeScholar
2023

PatchCT: Aligning Patch Set and Label Set with Conditional Transport for Multi-Label Image Classification

ICCV 2023poster

Multi-label image classification is a prediction task that aims to identify more than one label from a given image. This paper considers the semantic consistency of the latent space between the visual patch and linguistic label domains and introduces the conditional transport (CT) theory to bridge t…

Cited by 22PDFcodeScholar
2023

Tuning Multi-mode Token-level Prompt Alignment across Modalities

NeurIPS 2023poster

Advancements in prompt tuning of vision-language models have underscored their potential in enhancing open-world visual concept comprehension. However, prior works only primarily focus on single-mode (only one prompt for each modality) and holistic level (image or sentence) semantic alignment, which…

2022

Crossview Mapping with Graph-based Geolocalization on City-Scale Street Maps

ICRA 2022poster

3D environment mapping has been actively stud-ied recently with the development of autonomous driving and augmented reality. Although many image-based methods are proposed due to their convenience and flexibility compared to other complex sensors, few works focus on fixing the inherent scale ambigui…

Cited by 5SourceScholar
2022

DELTAR: Depth Estimation from a Light-Weight ToF Sensor and RGB Image

ECCV 2022poster

"Light-weight time-of-flight (ToF) depth sensors are small, cheap, low-energy and have been massively deployed on mobile devices for the purposes like autofocus, obstacle detection, etc. However, due to their specific measurements (depth distribution in a region instead of the depth value at a certa…