← Search

Ming Jiang

29 accepted papers

2025

No Preference Left Behind: Group Distributional Preference Optimization

ICLR 2025poster

Preferences within a group of people are not uniform but follow a distribution. While existing alignment methods like Direct Preference Optimization (DPO) attempt to steer models to reflect human preferences, they struggle to capture the distributional pluralistic preferences within a group. These m…

2025

SciEvent: Benchmarking Multi-domain Scientific Event Extraction

EMNLP 2025

Scientific information extraction (SciIE) has primarily relied on entity-relation extraction in narrow domains, limiting its applicability to interdisciplinary research and struggling to capture the necessary context of scientific information, often resulting in fragmented or conflicting statements.

2025

Towards Robust Few-Shot Relation Classification: Incorporating Relation Description with Agreement

EMNLP 2025

Few-shot relation classification aims to recognize the relation between two mentioned entities, with the help of only a few support samples. However, a few samples tend to be limited for tackling unlimited queries. If a query cannot find references from the support samples, it is defined as none-of-

2025

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

EMNLP 2025

Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mechanisms through input interventions. While textual interventions often target semantics, visual interventions typically rely on coarse pixel-level pertu

2024

Benchmarking Machine Translation with Cultural Awareness

EMNLP 2024finding

Translating culture-related content is vital for effective cross-cultural communication. However, many culture-specific items (CSIs) often lack literal translation across languages, making it challenging to collect high-quality, diverse parallel corpora with CSI annotations. This difficulty hinders…

2024

ECoK: Emotional Commonsense Knowledge Graph for Mining Emotional Gold

ACL 2024findings

The demand for understanding and expressing emotions in the field of natural language processing is growing rapidly. Knowledge graphs, as an important form of knowledge representation, have been widely utilized in various emotion-related tasks. However, existing knowledge graphs mainly focus on the…

2024

GRACE: Graph-Based Contextual Debiasing for Fair Visual Question Answering

ECCV 2024poster

"Large language models (LLMs) exhibit exceptional reasoning capabilities and have played significant roles in knowledge-based visual question-answering (VQA) systems. By conditioning on in-context examples and task-specific prompts, they comprehensively understand input questions and provide answers…

2024

GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths

ECCV 2024oral

"While exploring visual scenes, humans’ scanpaths are driven by their underlying attention processes. Understanding visual scanpaths is essential for various applications. Traditional scanpath models predict the where and when of gaze shifts without providing explanations, creating a gap in understa…

Cited by 4SourcePDFScholar
2024

Learning Chain of Counterfactual Thought for Bias-Robust Vision-Language Reasoning

ECCV 2024poster

"Despite the remarkable success of large vision-language models (LVLMs) on various tasks, their susceptibility to knowledge bias inherited from training data hinders their ability to generalize to new scenarios and limits their real-world applicability. To address this challenge, we propose the Coun…

2024

Lookahead Exploration with Neural Radiance Representation for Continuous Vision-Language Navigation

CVPR 2024highlight

Vision-and-language navigation (VLN) enables the agent to navigate to a remote location following the natural language instruction in 3D environments. At each navigation step the agent selects from possible candidate locations and then makes the move. For better navigation planning the lookahead exp…

2024

Simple but Effective Compound Geometric Operations for Temporal Knowledge Graph Completion

ACL 2024long

Temporal knowledge graph completion aims to infer the missing facts in temporal knowledge graphs. Current approaches usually embed factual knowledge into continuous vector space and apply geometric operations to learn potential patterns in temporal knowledge graphs. However, these methods only adopt…

2020

Fantastic Answers and Where to Find Them: Immersive Question-Directed Visual Attention

CVPR 2020poster

While most visual attention studies focus on bottom-up attention with restricted field-of-view, real-life situations are filled with embodied vision tasks. The role of attention is more significant in the latter due to the information overload, and attention to the most important regions is critical…

Cited by 23PDFScholar
2018

Beyond Trade-Off: Accelerate FCN-Based Face Detector With Higher Accuracy

CVPR 2018poster

Fully convolutional neural network (FCN) has been dominating the game of face detection task for a few years with its congenital capability of sliding-window-searching with shared kernels, which boiled down all the redundant calculation, and most recent state-of-the-art methods such as Faster-RCNN,…

Cited by 40SourcePDFScholar
2018

Emotional Attention: A Study of Image Sentiment and Visual Attention

CVPR 2018poster

Image sentiment influences visual perception. Emotion-eliciting stimuli such as happy faces and poisonous snakes are generally prioritized in human attention. However, little research has evaluated the interrelationships of image sentiment and visual saliency. In this paper, we present the first stu…

Cited by 187SourcePDFScholar
2016

A Paradigm for Building Generalized Models of Human Image Perception Through Data Fusion

CVPR 2016poster

In many sub-fields, researchers collect datasets of human ground truth that are used to create a new algorithm. For example, in research on image perception, datasets have been collected for topics such as what makes an image aesthetic or memorable. Despite high costs for human data collection, data…

Cited by 8PDFScholar