← Search

Hyun Joon Park

5 accepted papers

2025

Beyond Virtual Points: Depth-Enhanced LiDAR-only 3D Object Detection with Semi-Supervised Learning (Student Abstract)

AAAI 2025technical

The task of 3D object detection is crucial for various applications that rely on identifying objects in three-dimensional space using inputs like LiDAR point clouds and images. However, LiDAR-based detection faces challenges due to the sparsity of point clouds, especially at greater distances. To ad…

Cited by 0SourcePDFScholar
2023

AD-YOLO: You Look Only Once in Training Multiple Sound Event Localization and Detection

ICASSP 2023accepted

Sound event localization and detection (SELD) combines the identification of sound events with the corresponding directions of arrival (DOA). Recently, event-oriented track output formats have been adopted to solve this problem; however, they still have limited generalization toward real-world probl…

Cited by 0SourceScholar
2023

MetricGAN-OKD: Multi-Metric Optimization of MetricGAN via Online Knowledge Distillation for Speech Enhancement

ICML 2023poster

In speech enhancement, MetricGAN-based approaches reduce the discrepancy between the $L_p$ loss and evaluation metrics by utilizing a non-differentiable evaluation metric as the objective function. However, optimizing multiple metrics simultaneously remains challenging owing to the problem of confus…

Cited by 9SourcePDFScholar
2023

TriAAN-VC: Triple Adaptive Attention Normalization for Any-to-Any Voice Conversion

ICASSP 2023accepted

Voice Conversion (VC) must be achieved while maintaining the content of the source speech and representing the characteristics of the target speaker. The existing methods do not simultaneously satisfy the above two aspects of VC, and their conversion outputs suffer from a trade-off problem between m…

Cited by 0SourceScholar
2022

MANNER: Multi-View Attention Network For Noise Erasure

ICASSP 2022accepted

In the field of speech enhancement, time domain methods have difficulties in achieving both high performance and efficiency. Recently, dual-path models have been adopted to represent long sequential features, but they still have limited representations and poor memory efficiency. In this study, we p…

Cited by 0SourceScholar