← Search

Fei Xie

13 accepted papers

2026

AutoFly: Vision-Language-Action Model for UAV Autonomous Navigation in the Wild

ICLR 2026poster

Vision-language navigation (VLN) requires intelligent agents to navigate environments by interpreting linguistic instructions alongside visual observations, serving as a cornerstone task in Embodied AI. Current VLN research for unmanned aerial vehicles (UAVs) relies on detailed, pre-specified instru…

Cited by 0SourceScholar
2026

DirectFisheye-GS: Enabling Native Fisheye Input in Gaussian Splatting with Cross-View Joint Optimization

CVPR 2026

3D Gaussian Splatting (3DGS) has enabled efficient 3D scene reconstruction from everyday images with real-time, high-fidelity rendering, greatly advancing VR/AR applications. Fisheye cameras, with their wider field of view (FOV), promise high-quality reconstructions from fewer inputs and have recent

Cited by 1SourceScholar
2026

JetsonCompletion: Real-Time Depth Completion on Resource-Constrained Edge Devices

ICRA 2026poster

Depth completion from sparse LiDAR points and images is a key perception task for autonomous robots, enabling dense 3D understanding in challenging environments. However, most recent researches achieve accuracy gains by greatly enlarging network size, making them unsuitable for realtime deployment o…

Cited by 0codeScholar
2025

Mamba-Adaptor: State Space Model Adaptor for Visual Recognition

CVPR 2025poster

Recent State Space Models (SSM), especially Mamba, have demonstrated impressive performance in visual modeling and possess superior model efficiency. However, the application of Mamba to visual tasks suffers inferior performance due to three main constraints existing in the sequential model: 1) Casu…

Cited by 0SourcePDFScholar
2025

PVMamba: Parallelizing Vision Mamba via Dynamic State Aggregation

ICCV 2025poster

Mamba, an architecture with RNN-like sequence modeling of State Space Model (SSM), has demonstrated promising capabilities in long-range modeling with high efficiency. However, Mamba models struggle with structured 2D visual data using sequential computing, thereby lagging behind their attention-bas…

2024

QuadMamba: Learning Quadtree-based Selective Scan for Visual State Space Model

NeurIPS 2024poster

Recent advancements in State Space Models, notably Mamba, have demonstrated superior performance over the dominant Transformer models, particularly in reducing the computational complexity from quadratic to linear. Yet, difficulties in adapting Mamba from language to vision tasks arise due to the di…

2024

Towards Category Unification of 3D Single Object Tracking on Point Clouds

ICLR 2024poster

Category-specific models are provenly valuable methods in 3D single object tracking (SOT) regardless of Siamese or motion-centric paradigms. However, such over-specialized model designs incur redundant parameters, thus limiting the broader applicability of 3D SOT task. This paper first introduces un…

Cited by 12SourcePDFScholar
2020

Adaptive Cross-Coupled Control of Cable-Driven Parallel Robots With Model Uncertainties

RA-L 2020

Cable-driven parallel robots (CDPRs) are robots with novel structures, wherein flexible cables, instead of rigid links, are employed to pull mobile platforms. This structural change enables CDPRs to not only offer potential advantages, but also introduces control challenges with regard to frictional

Cited by 37SourceScholar
2018

Deep Stock Representation Learning: From Candlestick Charts to Investment Decisions

ICASSP 2018accepted

We propose a novel investment decision strategy (IDS) based on deep learning. The performance of many IDSs is affected by stock similarity. Most existing stock similarity measurements have the problems: (a) The linear nature of many measurements cannot capture nonlinear stock dynamics; (b) The estim…

Cited by 0SourceScholar