← Search

Jihwan Kim

12 accepted papers

2026

LiveWeb-IE: A Benchmark For Online Web Information Extraction

ICLR 2026poster

Web information extraction (WIE) is the task of automatically extracting data from web pages, offering high utility for various applications. The evaluation of WIE systems has traditionally relied on benchmarks built from HTML snapshots captured at a single point in time. However, this offline evalu…

Cited by 0SourceScholar
2025

Prediction-Feedback DETR for Temporal Action Detection

AAAI 2025technical

Temporal Action Detection (TAD) is fundamental yet challenging for real-world video applications. Leveraging the unique benefits of transformers, various DETR-based approaches have been adopted in TAD. However, it has recently been identified that the attention collapse in self-attention causes the…

Cited by 1SourcePDFScholar
2024

EquiGraspFlow: SE(3)-Equivariant 6-DoF Grasp Pose Generative Flows

CoRL 2024poster

Traditional methods for synthesizing 6-DoF grasp poses from 3D observations often rely on geometric heuristics, resulting in poor generalizability, limited grasp options, and higher failure rates. Recently, data-driven methods have been proposed that use generative models to learn the distribution o…

Cited by 7SourcecodeScholar
2024

FIFO-Diffusion: Generating Infinite Videos from Text without Training

NeurIPS 2024poster

We propose a novel inference technique based on a pretrained diffusion model for text-conditional video generation. Our approach, called FIFO-Diffusion, is conceptually capable of generating infinitely long videos without additional training. This is achieved by iteratively performing diagonal denoi…

2024

Graph Geometry-Preserving Autoencoders

ICML 2024poster

When using an autoencoder to learn the low-dimensional manifold of high-dimensional data, it is crucial to find the latent representations that preserve the geometry of the data manifold. However, most existing studies assume a Euclidean nature for the high-dimensional data space, which is arbitrary…

2021

Learning-Based Real-Time Detection of Robot Collisions Without Joint Torque Sensors

RA-L 2021

Robots operating in close proximity to humans require fast and reliable detection of collisions, which can range from sharp impacts (hard collisions) to pulling-pushing-catching motions (soft collisions). Because joint torque sensors can be costly, the external joint torques caused by collisions are

Cited by 73SourceScholar
2021

Self-Supervised Video GANs: Learning for Appearance Consistency and Motion Coherency

CVPR 2021poster

A video can be represented by the composition of appearance and motion. Appearance (or content) expresses the information invariant throughout time, and motion describes the time-variant movement. Here, we propose self-supervised approaches for video Generative Adversarial Networks (GANs) to achieve…

Cited by 25PDFScholar