← Search

YI Shen

18 accepted papers

2026

JRDB-Reasoning: A Difficulty-Graded Benchmark for Visual Reasoning in Robotics

AAAI 2026technical

Recent advances in Vision-Language Models (VLMs) and large language models (LLMs) have greatly enhanced visual reasoning, a key capability for embodied AI agents like robots. However, existing visual reasoning benchmarks often suffer from several limitations: they lack a clear definition of reasonin

Cited by 0SourcePDFScholar
2026

LD-MoLE: Learnable Dynamic Routing for Mixture of LoRA Experts

ICLR 2026poster

Recent studies have shown that combining parameter-efficient fine-tuning (PEFT) with mixture-of-experts (MoE) is an effective strategy for adapting large language models (LLMs) to the downstream tasks. However, most existing approaches rely on conventional TopK routing, which requires careful hyperp…

Cited by 0SourcecodeScholar
2025

Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration

ICML 2025poster

Learning to cooperate in distributed partially observable environments with no communication abilities poses significant challenges for multi-agent deep reinforcement learning (MARL). This paper addresses key concerns in this domain, focusing on inferring state representations from individual agent…

2025

Improve Decoding Factuality by Token-wise Cross Layer Entropy of Large Language Models

NAACL 2025findings

Despite their impressive capacities, Large language models (LLMs) often struggle with the hallucination issue of generating inaccurate or fabricated content even when they possess correct knowledge. In this paper, we extend the exploration of the correlation between hidden-state prediction changes a…

Cited by 0SourcePDFScholar
2025

YOLO-MARL: You Only LLM Once for Multi-Agent Reinforcement Learning

IROS 2025

Advancements in deep multi-agent reinforcement learning (MARL) have positioned it as a promising approach for decision-making in cooperative games. However, it still remains challenging for MARL agents to learn cooperative strategies for some game environments. Recently, large language models (LLMs)

Cited by 8SourcecodeScholar
2024

Online Learning Based Shape Control for a Soft Manipulator Based on Spatial Features Feedback

RA-L 2024

Although soft manipulators are endowed with compliance and flexibility, most control strategies focus on end-effector control and lack shape control ability. This letter aims to design a shape controller for the soft manipulator. Firstly, we establish a modified forward kinematics model (FKM) based

Cited by 5SourceScholar
2024

Outlier-Robust Distributionally Robust Optimization via Unbalanced Optimal Transport

NeurIPS 2024poster

Distributionally Robust Optimization (DRO) accounts for uncertainty in data distributions by optimizing the model performance against the worst possible distribution within an ambiguity set. In this paper, we propose a DRO framework that relies on a new distance inspired by Unbalanced Optimal Transp…

Cited by 14SourcePDFScholar
2022

Design of a Soft Gripper With Improved Microfluidic Tactile Sensors for Classification of Deformable Objects

RA-L 2022

Tactile object recognition is vital for robotic handling systems; however, existing technologies that concentrate on tactile sensors with high modulus are not suitable for soft grippers to classify deformable objects. In this letter, we integrated an indenter layer into the traditional microfluidic

Cited by 19SourceScholar
2022

Sen-Glove: A Lightweight Wearable Glove for Hand Assistance with Soft Joint Sensing

ICRA 2022poster

Perception and portability are critical issues for wearable gloves in hand assistive engineering. However, available wearable gloves either lack flexible sensing or are bulky. In this paper, we present a tendon-driven lightweight wearable glove with soft joint sensing, Sen-Glove. Sen-Glove is equipp…

Cited by 12SourceScholar
2022

Seq2Path: Generating Sentiment Tuples as Paths of a Tree

ACL 2022findings

Aspect-based sentiment analysis (ABSA) tasks aim to extract sentiment tuples from a sentence. Recent generative methods such as Seq2Seq models have achieved good performance by formulating the output as a sequence of sentiment tuples. However, the orders between the sentiment tuples do not naturally…

Cited by 85SourcePDFScholar
2021

A Joint Training Dual-MRC Framework for Aspect Based Sentiment Analysis

AAAI 2021technical

Aspect based sentiment analysis (ABSA) involves three fundamental subtasks: aspect term extraction, opinion term extraction, and aspect-level sentiment classification. Early works only focused on solving one of these subtasks individually. Some recent work focused on solving a combination of two sub…

2020

VectorNet: Encoding HD Maps and Agent Dynamics From Vectorized Representation

CVPR 2020poster

Behavior prediction in dynamic, multi-agent systems is an important problem in the context of self-driving cars, due to the complex representations and interactions of road components, including moving agents (e.g. pedestrians and vehicles) and road context information (e.g. lanes, traffic lights).…

Cited by 1022PDFScholar
2019

A Hierarchical Framework for Coordinating Large-Scale Robot Networks

ICRA 2019poster

In this paper, we study the cooperative path planning and motion coordination problems of the multi-robot system with large number of robots, aiming for practical applications in robotic warehouses and automated transportation systems. Particularly, we solve the life-long planning problem and guaran…

Cited by 15SourceScholar
2019

Objective Comparison of Speech Enhancement Algorithms with Hearing Loss Simulation

ICASSP 2019accepted

Many speech enhancement algorithms have been proposed over the years and it has been shown that deep neural networks can lead to significant improvements. These algorithms, however, have not been validated for hearing-impaired listeners. Additionally, these algorithms are often evaluated under a lim…

Cited by 8SourceScholar