← Search

Jana Kosecka

6 accepted papers

2024

Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing

EMNLP 2024finding

Sign language translation from video to spoken text presents unique challenges owing to the distinct grammar, expression nuances, and high variation of visual appearance across different speakers and contexts. Gloss annotations serve as an intermediary to guide the translation process. In our work,…

2019

Learning Local RGB-to-CAD Correspondences for Object Pose Estimation

ICCV 2019poster

We consider the problem of 3D object pose estimation. While much recent work has focused on the RGB domain, the reliance on accurately annotated images limits generalizability and scalability. On the other hand, the easily available object CAD models are rich sources of data, providing a large numbe…

Cited by 30PDFScholar
2017

3D Bounding Box Estimation Using Deep Learning and Geometry

CVPR 2017poster

We present a method for 3D object detection and pose estimation from a single image. In contrast to current techniques that only regress the 3D orientation of an object, our method first regresses relatively stable 3D object properties using a deep convolutional neural network and then combines thes…

Cited by 1362PDFScholar
2017

Synthesizing Training Data for Object Detection in Indoor Scenes

RSS 2017poster

Detection of objects in cluttered indoor environments is one of the key enabling functionalities for service robots. The best performing object detection approaches in computer vision exploit deep Convolutional Neural Networks (CNN) to simultaneously detect and categorize the objects of interest in…

Cited by 285SourcePDFScholar