← Search

Namkoo Ha

2 accepted papers

2025

First Attentions Last: Better Exploiting First Attentions for Efficient Parallel Training

NeurIPS 2025poster

As training billion-scale transformers becomes increasingly common, employing multiple distributed GPUs along with parallel training methods has become a standard practice. However, existing transformer designs suffer from significant communication overhead, especially in Tensor Parallelism (TP), wh…

Cited by 0SourceScholar
2021

Deep Low-Contrast Image Enhancement using Structure Tensor Representation

AAAI 2021technical

We present a new deep learning framework for low-contrast image enhancement, which trains the network using the multi-exposure sequences rather than explicit ground-truth images. The purpose of our method is to enhance a low-contrast image so as to contain abundant details in various exposure levels…

Cited by 2SourcePDFScholar