← Search

Qichao Ma

6 accepted papers

2026

GLARE: Scalable Neuro-Symbolic Reward Shaping for LLM Agents via Group-Level Automata

ICML 2026poster

Reinforcement Learning (RL) with Group Relative Policy Optimization (GRPO) shows great promise for enhancing LLM reasoning, but remains challenged by sparse and unstable rewards in long-horizon tasks. Existing approaches to reward shaping struggle to balance semantic expressiveness, reliability, and…

Cited by 0SourceScholar
2025

Faster and Stronger: When ANN-SNN Conversion Meets Parallel Spiking Calculation

ICML 2025poster

Spiking Neural Network (SNN), as a brain-inspired and energy-efficient network, is currently facing the pivotal challenge of exploring a suitable and efficient learning framework. The predominant training methodologies, namely Spatial-Temporal Back-propagation (STBP) and ANN-SNN Conversion, are encu…

2024

A Non-Homogeneity Mapless Navigation Based on Hierarchical Safe Reinforcement Learning in Dynamic Complex Environments

IROS 2024poster

Addressing safe and efficient navigation in dynamic, realistic, and complex environments stands as a pivotal inquiry within the realm of robotics. Recently, numerous learning-based methods are introduced into the field of navigation, yielding notable outcomes. In this letter, we propose a hierarchic…

Cited by 0SourceScholar
2024

PE-Planner: A Performance-Enhanced Quadrotor Motion Planner for Autonomous Flight in Complex and Dynamic Environments

RA-L 2024

The role of a motion planner is pivotal in quadrotor applications, yet existing methods often struggle to adapt to complex environments, limiting their ability to achieve fast, safe, and robust flight. In this letter, we introduce a performance-enhanced quadrotor motion planner designed for autonomo

Cited by 10SourceScholar
2024

SRL-ORCA: A Socially Aware Multi-Agent Mapless Navigation Algorithm in Complex Dynamic Scenes

RA-L 2024

For real-world navigation, it is important to endow robots with the capabilities to navigate safely and efficiently in a complex environment with both dynamic and static obstacles. However, achieving path-finding in non-convex complex environments without maps as well as enabling multiple robots to

Cited by 22SourceScholar
2023

Image-Based Visual Servoing of Quadrotors to Arbitrary Flight Targets

RA-L 2023

Visual servoing of Unmanned Aerial Vehicles (UAVs) has achieved satisfactory performance in fixed and planar motion targets. Due to highly coupled system dynamics and the sensitivity of the target image to aircraft attitude, the problem for chasing free-flying targets remains challenging. In this pa

Cited by 22SourceScholar