← Search

Yifeng Li

8 accepted papers

2025

Can We Afford The Perfect Prompt? Balancing Cost and Accuracy with the Economical Prompting Index

COLING 2025main

As prompt engineering research rapidly evolves, evaluations beyond accuracy are crucial for developing cost-effective techniques. We present the Economical Prompting Index (EPI), a novel metric that combines accuracy scores with token consumption, adjusted by a user-specified cost concern level to r…

2025

Unveiling Local Well-posedness Influence for Cross-modal Person Re-Identification

ICASSP 2025accepted

The existing cross-modal retrieval methods trend toward the conventional multi-modal alignment while ignoring the localization bias caused by visual hallucination, including color pollution and appearance-like occlusion due to uncontrollable factors such as weather, illumination, and occlusion. This…

Cited by 0SourceScholar
2024

Picturing Ambiguity: A Visual Twist on the Winograd Schema Challenge

ACL 2024long

Large Language Models (LLMs) have demonstrated remarkable success in tasks like the Winograd Schema Challenge (WSC), showcasing advanced textual common-sense reasoning. However, applying this reasoning to multimodal domains, where understanding text and images together is essential, remains a substa…

2024

Towards Unified Interactive Visual Grounding in The Wild

ICRA 2024poster

Interactive visual grounding in Human-Robot Interaction (HRI) is challenging yet practical due to the inevitable ambiguity in natural languages. It requires robots to disambiguate the user’s input by active information gathering. Previous approaches often rely on predefined templates to ask disambig…

Cited by 3SourcecodeScholar
2022

Generative Category-Level Shape and Pose Estimation with Semantic Primitives

CoRL 2022poster

Empowering autonomous agents with 3D understanding for daily objects is a grand challenge in robotics applications. When exploring in an unknown environment, existing methods for object pose estimation are still not satisfactory due to the diversity of object shapes. In this paper, we propose a nove…

Cited by 28SourcecodeScholar
2021

Simultaneous Semantic and Collision Learning for 6-DoF Grasp Pose Estimation

IROS 2021poster

Grasping in cluttered scenes has always been a great challenge for robots, due to the requirement of the ability to well understand the scene and object information. Previous works usually assume that the geometry information of the objects is available, or utilize a step-wise, multi-stage strategy…

Cited by 64SourcecodeScholar
2020

Intra-Correlation Encoding for Chinese Sentence Intention Matching

COLING 2020main

Sentence intention matching is vital for natural language understanding. Especially for Chinese sentence intention matching task, due to the ambiguity of Chinese words, semantic missing or semantic confusion are more likely to occur in the encoding process. Although the existing methods have enriche…