← Search

Katsushi Ikeuchi

9 accepted papers

2024

GPT-4V(ision) for Robotics: Multimodal Task Planning From Human Demonstration

RA-L 2024

We introduce a pipeline that enhances a general-purpose Vision Language Model, GPT-4V(ision), to facilitate one-shot visual teaching for robotic manipulation. This system analyzes videos of humans performing tasks and outputs executable robot programs that incorporate insights into affordances. The

Cited by 117SourceScholar
2024

LiDAR-camera Calibration using Intensity Variance Cost

ICRA 2024poster

We propose an extrinsic calibration method for LiDAR-camera fusion systems using variations in intensities projected from camera images to the LiDAR point cloud. As the input, the proposed method uses a sequence of LiDAR data and camera images captured while moving the system. Once the camera motion…

Cited by 2SourceScholar
2022

Deep Gesture Generation for Social Robots Using Type-Specific Libraries

IROS 2022poster

Body language such as conversational gesture is a powerful way to ease communication. Conversational gestures do not only make a speech more lively but also contain semantic meaning that helps to stress important information in the discussion. In the field of robotics, giving conversational agents (…

Cited by 6SourceScholar
2021

Task-Oriented Motion Mapping on Robots of Various Configuration Using Body Role Division

RA-L 2021

Many works in robot teaching either focus only on teaching task knowledge, such as geometric constraints, or motion knowledge, such as the motion for accomplishing a task. However, to effectively teach a complex task sequence to a robot, it is important to take advantage of both task and motion know

Cited by 24SourceScholar
2021

Verbal Focus-of-Attention System for Learning-from-Observation

ICRA 2021poster

The learning-from-observation (LfO) framework aims to map human demonstrations to a robot to reduce programming effort. To this end, an LfO system encodes a human demonstration into a series of execution units for a robot, which are referred to as task models. Although previous research has proposed…

Cited by 17SourceScholar
2018

LiDAR and Camera Calibration Using Motions Estimated by Sensor Fusion Odometry

IROS 2018poster

This paper proposes a targetless and automatic camera-LiDAR calibration method. Our approach extends the hand-eye calibration framework to 2D-3D calibration. The scaled camera motions are accurately calculated using a sensor-fusion odometry method. We also clarify the suitable motions for our calibr…

Cited by 168SourceScholar
2016

Reconstructing Shapes and Appearances of Thin Film Objects Using RGB Images

CVPR 2016poster

Reconstruction of shapes and appearances of thin film objects can be applied to many fields such as industrial inspection, biological analysis, and archeology research. However, it comes with many challenging issues because the appearances of thin film can change dramatically depending on view and l…

Cited by 10PDFScholar