← Search

Song Li

16 accepted papers

2026

MemOcc: Hierarchical Memory for Indoor Continuous Occupancy Mapping

ICRA 2026poster

Indoor 3D occupancy mapping, crucial for robotic perception, struggles with occlusions and reappearing surfaces in continuous observations. Existing methods either fuse frames without discernment, causing occlusion-induced errors to persist and contaminate global representations, or recompute scenes…

Cited by 0Scholar
2025

CogMath: Assessing LLMs' Authentic Mathematical Ability from a Human Cognitive Perspective

ICML 2025poster

Although large language models (LLMs) show promise in solving complex mathematical tasks, existing evaluation paradigms rely solely on a coarse measure of overall answer accuracy, which are insufficient for assessing their authentic capabilities. In this paper, we propose \textbf{CogMath}, which com…

Cited by 0SourcePDFScholar
2025

Immunogenicity Prediction with Dual Attention Enables Vaccine Target Selection

ICLR 2025poster

Immunogenicity prediction is a central topic in reverse vaccinology for finding candidate vaccines that can trigger protective immune responses. Existing approaches typically rely on highly compressed features and simple model architectures, leading to limited prediction accuracy and poor generaliza…

2024

A Dragonfly-inspired Flapping Wing Robot Mimicking Force Vector Control Approach

ICRA 2024poster

Dragonflies show impressive flying skills by achieving both high efficiency and agility. They can perform distinctive flight maneuvers, such as flying backwards, which has proven to be achieved through "force vectoring" mechanism recently. In this paper, to explore the agile flight ability of dragon…

Cited by 3SourceScholar
2024

Behavioral Recognition of Skeletal Data Based on Targeted Dual Fusion Strategy

AAAI 2024technical

The deployment of multi-stream fusion strategy on behavioral recognition from skeletal data can extract complementary features from different information streams and improve the recognition accuracy, but suffers from high model complexity and a large number of parameters. Besides, existing multi-str…

2024

Enhancing Multilingual Speech Recognition through Language Prompt Tuning and Frame-Level Language Adapter

ICASSP 2024accepted

Multilingual intelligent assistants, such as ChatGPT, have recently gained popularity. To further expand the applications of multilingual artificial intelligence (AI) assistants and facilitate international communication, it is essential to enhance the performance of multilingual speech recognition,…

Cited by 0SourceScholar
2024

Virtual Scanning: Unsupervised Non-line-of-sight Imaging from Irregularly Undersampled Transients

NeurIPS 2024poster

Non-line-of-sight (NLOS) imaging allows for seeing hidden scenes around corners through active sensing. Most previous algorithms for NLOS reconstruction require dense transients acquired through regular scans over a large relay surface, which limits their applicability in realistic scenarios with ir…

2023

AirTwins: Modular Bi-Copters Capable of Splitting From Their Combined Quadcopter in Midair

RA-L 2023

Micro tandem bi-copters are capable of passing through narrow gaps owing to their particular slender shape. However, the introduction of the tilting servo motors leads to a non-minimum phase roll dynamics, which affects their flight stability when exploring environments with unpredictable disturbanc

Cited by 14SourceScholar
2022

Liftoff of A Motor-Driven Flapping Wing Rotorcraft with Mechanically Decoupled Wings

ICRA 2022poster

Flapping Wing Rotorcraft (FWR) combines flapping and rotating wing motion in one element. Such a hybrid design integrates the high-efficiency characteristics of the rotating wing and the high-lift feature of the flapping wing under low Reynolds number, providing a broader range of simultaneous lift…

Cited by 7SourceScholar
2021

End-To-End Multi-Accent Speech Recognition with Unsupervised Accent Modelling

ICASSP 2021accepted

End-to-end speech recognition has achieved good recognition performance on standard English pronunciation datasets. However, one prominent problem with end-to-end speech recognition systems is that non-native English speakers tend to have complex and varied accents, which reduces the accuracy of Eng…

Cited by 0SourceScholar
2020

Cartoon-Texture Decomposition-Based Variational Pansharpening

ICASSP 2020accepted

Pansharpening is widely used to increase the spatial resolution of a multispectral (MS) image by fusing with a panchromatic (PAN) image that has high-spatial resolution and the same scene. In this paper, the similarities of MS and PAN images in cartoon-texture space are exploited. The cartoon and te…

Cited by 0SourceScholar
2020

Crop Height and Plot Estimation for Phenotyping from Unmanned Aerial Vehicles using 3D LiDAR

IROS 2020poster

We present techniques to measure crop heights using a 3D Light Detection and Ranging (LiDAR) sensor mounted on an Unmanned Aerial Vehicle (UAV). Knowing the height of plants is crucial to monitor their overall health and growth cycles, especially for high-throughput plant phenotyping. We present a m…

Cited by 26SourcecodeScholar
2019

A Hybrid Method for Blind Estimation of Frequency Dependent Reverberation Time Using Speech Signals

ICASSP 2019accepted

Reverberation time is an important room acoustical parameter that can be used to identify the acoustic environment, predict speech intelligibility and model the late reverberation for binaural rendering, etc. Several blind estimation algorithms of reverberation time have been proposed by analyzing r…

Cited by 0SourceScholar