← Search

Jongseok Kim

6 accepted papers

2024

IncepSeqNet: Advancing Signal Classification with Multi-Shape Augmentation (Student Abstract)

AAAI 2024technical

This work proposes and analyzes IncepSeqNet which is a new model combining the Inception Module with the innovative Multi-Shape Augmentation technique. IncepSeqNet excels in feature extraction from sequence signal data consisting of a number of complex numbers to achieve superior classification accu…

Cited by 4SourcePDFScholar
2021

Dual Compositional Learning in Interactive Image Retrieval

AAAI 2021technical

We present an approach named Dual Composition Network (DCNet) for interactive image retrieval that searches for the best target image for a natural language query and a reference image. To accomplish this task, existing methods have focused on learning a composite representation of the reference ima…

Cited by 103SourcePDFScholar
2021

Transitional Adaptation of Pretrained Models for Visual Storytelling

CVPR 2021poster

Previous models for vision-to-language generation tasks usually pretrain a visual encoder and a language generator in the respective domains and jointly finetune them with the target task. However, this direct transfer practice may suffer from the discord between visual specificity and language flue…

Cited by 40PDFcodeScholar
2021

Viewpoint-Agnostic Change Captioning With Cycle Consistency

ICCV 2021poster

Change captioning is the task of identifying the change and describing it with a concise caption. Despite recent advancements, filtering out insignificant changes still remains as a challenge. Namely, images from different camera perspectives can cause issues; a mere change in viewpoint should be di…

Cited by 44PDFcodeScholar
2020

Character Grounding and Re-Identification in Story of Videos and Text Descriptions

ECCV 2020poster

We address character grounding and re-identification in multiple story-based videos like movies and associated text descriptions. In order to solve these related tasks in a mutually rewarding way, we propose a model named Character in Story Identification Network (CiSIN). Our method builds two seman…