← Search

Hui Song

5 accepted papers

2025

Boosting Text-to-SQL through Multi-grained Error Identification

COLING 2025main

Text-to-SQL is a technology that converts natural language questions into executable SQL queries, allowing users to query and manage relational databases more easily. In recent years, large language models have significantly advanced the development of text-to-SQL. However, existing methods often ov…

2025

Enhancing Multimodal Named Entity Recognition through Adaptive Mixup Image Augmentation

COLING 2025main

Multimodal named entity recognition (MNER) extends traditional named entity recognition (NER) by integrating visual and textual information. However, current methods still face significant challenges due to the text-image mismatch problem. Recent advancements in text-to-image synthesis provide promi…

Cited by 0SourcePDFScholar
2023

Deep Learning-Based Path Loss Prediction for Outdoor Wireless Communication Systems

ICASSP 2023accepted

Deep learning (DL) has been recently leveraged for the inference of characteristics related to wireless communication channels, such as path loss (PL). This paper presents how a deep convolutional encoder-decoder, namely a path loss prediction net (PPNet) based on SegNet, can be trained to transform…

Cited by 0SourceScholar
2022

Different Data, Different Modalities! Reinforced Data Splitting for Effective Multimodal Information Extraction from Social Media Posts

COLING 2022main

Recently, multimodal information extraction from social media posts has gained increasing attention in the natural language processing community. Despite their success, current approaches overestimate the significance of images. In this paper, we argue that different social media posts should consid…

2020

DNN-based Mask Estimation Integrating Spectral and Spatial Features for Robust Beamforming

ICASSP 2020accepted

Spectral mask based beamforming has showed competitive performance on multi-channel speech enhancement in recent years. However, such methods apply mask estimation on each channel and ensemble the masks from multiple channels into one for speech and noise covariance estimation. Spectral-spatial mask…

Cited by 0SourceScholar