← Search

Ehsan Hoque

5 accepted papers

2026

OmniSapiens: A Foundation Model for Social Behavior Processing via Heterogeneity-Aware Relative Policy Optimization

ICML 2026poster

To develop socially intelligent AI, existing approaches typically model behavioral dimensions (e.g., affective, cognitive, or social attributes) in isolation. Although useful, this task-specific modeling increases training costs and limits generalization across behavioral settings. Recent reasoning …

Cited by 0SourceScholar
2025

Accessible, At-Home Detection of Parkinson’s Disease via Multi-Task Video Analysis

AAAI 2025technical

Limited accessibility to neurological care leads to under-diagnosed Parkinson's Disease (PD), preventing early intervention. Existing AI-based PD detection methods primarily focus on unimodal analysis of motor or speech tasks, overlooking the multifaceted nature of the disease. To address this, we i…

2024

A Survey on Open Information Extraction from Rule-based Model to Large Language Model

EMNLP 2024finding

Open Information Extraction (OpenIE) represents a crucial NLP task aimed at deriving structured information from unstructured text, unrestricted by relation type or domain. This survey paper provides an overview of OpenIE technologies spanning from 2007 to 2024, emphasizing a chronological perspecti…

Cited by 3SourcePDFScholar
2021

Hitting your MARQ: Multimodal ARgument Quality Assessment in Long Debate Video

EMNLP 2021main

The combination of gestures, intonations, and textual content plays a key role in argument delivery. However, the current literature mostly considers textual content while assessing the quality of an argument, and it is limited to datasets containing short sequences (18-48 words). In this paper, we…

Cited by 5SourcePDFScholar
2021

Humor Knowledge Enriched Transformer for Understanding Multimodal Humor

AAAI 2021technical

Recognizing humor from a video utterance requires understanding the verbal and non-verbal components as well as incorporating the appropriate context and external knowledge. In this paper, we propose Humor Knowledge enriched Transformer (HKT) that can capture the gist of a multimodal humorous expres…