← Search

Caigui Jiang

4 accepted papers

2026

W2S-AlignTree: Weak-to-Strong Inference-Time Alignment for Large Language Models via Monte Carlo Tree Search

AAAI 2026technical

Large Language Models (LLMs) demonstrate impressive capabilities, yet their outputs often suffer from misalignment with human preferences due to the inadequacy of weak supervision and a lack of fine-grained control. Training-time alignment methods like Reinforcement Learning from Human Feedback (RLH

Cited by 0SourcePDFScholar
2025

AGCNet: Improving Inertial Odometry via IMU Accelerometer and Gyroscope Online Compensation

IROS 2025

This paper presents a learning-based online IMU compensation method (AGCNet) that can compensate for run-time errors of the accelerometer and gyroscope to improve inertial odometry. AGCNet employs U-Net architecture with hybrid dilated convolutions to extract multiscale features. It also adopts skip

Cited by 0SourceScholar
2024

GSO-Net: Grid Surface Optimization via Learning Geometric Constraints

AAAI 2024technical

In the context of surface representations, we find a natural structural similarity between grid surface and image data. Motivated by this inspiration, we propose a novel approach: encoding grid surfaces as geometric images and using image processing methods to address surface optimization-related pr…