← Search

Jiahao Huang

7 accepted papers

2026

GEMA-Score: Granular Explainable Multi-Agent Scoring Framework for Radiology Report Evaluation

AAAI 2026technical

Automatic medical report generation has the potential to support clinical diagnosis, reduce the workload of radiologists, and demonstrate potential for enhancing diagnostic consistency. However, current evaluation metrics often fail to reflect the clinical reliability of generated reports. Overlap-b

Cited by 0SourcePDFScholar
2026

Nano-EmoX: Unifying Multimodal Emotional Intelligence from Perception to Empathy

CVPR 2026

The development of affective multimodal language models (MLMs) has long been constrained by a gap between low-level perception and high-level interaction, leading to fragmented affective capabilities and limited generalization. To bridge this gap, we propose a cognitively inspired three-level hierar

Cited by 0SourceScholar
2025

JMedBench: A Benchmark for Evaluating Japanese Biomedical Large Language Models

COLING 2025main

Recent developments in Japanese large language models (LLMs) primarily focus on general domains, with fewer advancements in Japanese biomedical LLMs. One obstacle is the absence of a comprehensive, large-scale benchmark for comparison. Furthermore, the resources for evaluating Japanese biomedical LL…

2025

Leveraging High-Resource English Corpora for Cross-lingual Domain Adaptation in Low-Resource Japanese Medicine via Continued Pre-training

EMNLP 2025

Limited low-resource language corpora in professional domains like medicine hinder cross-lingual domain adaptation of pre-trained large language models (PLMs). While abundant English medical corpora could complement this scarcity, the effective mixture of English and target language, including machi

2025

Narrowing Information Bottleneck Theory for Multimodal Image-Text Representations Interpretability

ICLR 2025poster

The task of identifying multimodal image-text representations has garnered increasing attention, particularly with models such as CLIP (Contrastive Language-Image Pretraining), which demonstrate exceptional performance in learning complex associations between images and text. Despite these advanceme…

2025

TactfulToM: Do LLMs have the Theory of Mind ability to understand White Lies?

EMNLP 2025

While recent studies explore Large Language Models’ (LLMs) performance on Theory of Mind (ToM) reasoning tasks, research on ToM abilities that require more nuanced social context is limited, such as white lies. We introduce TactfulToM, a novel English benchmark designed to evaluate LLMs’ ability to

2024

HAMLET: Graph Transformer Neural Operator for Partial Differential Equations

ICML 2024poster

We present a novel graph transformer framework, HAMLET, designed to address the challenges in solving partial differential equations (PDEs) using neural networks. The framework uses graph transformers with modular input encoders to directly incorporate differential equation information into the solu…

Cited by 10SourcePDFScholar